I had two, then only one once the liquor store down road closed up shop.
HN user
thelastparadise
But do we know that the same techniques won't scale if trained in the huge clusters?
The megga disappointment is o1 is performing worse than o1-preview [1], and claude 3.6 had already nearly caught up to o1-preview.
I could a car for that kind of money!
Is it working?
Do they denied their surgery or did they get it but bankrupt after?
What do you should be done with him then?
So what you're saying is Intel, or any other would-be NVIDIA competitor, needs to put out fast interconnects, not just compute cards. This is true.
I'm not sure your argument stands when it comes to OP's idea of a single card with 128GB VRAM. This would be enough to run ~180B models with reasonable quantization --we're not near maxing out the capability of 180B yet (see the latest 32B models performing near public SOTA).
This indeed would push rapid and wide adoption and be quite disruptive. But sure, it wouldn't instantly enable competitive training of 405B models.
Thank you.
No one ever talks about just how easy it is to add GPIO to anything.
Yeah, this is is true. I've been taking for 3 months now --best consistent sleep of my life, but definitely occasionally the most terrifying sleep paralysis I've ever experienced, or even heard of.
I know it sounds ridiculous, but ~1-2 nights of absolute terror per week is totally worth it compared to how it used to be (getting maybe ~1 good night of sleep every couple weeks.)
If they did that to me I would take it as a sign of disrespect and resign immediately.
Oops now you gotta hire a new one.
Several cases like this on here. These would be valid wrongful termination lawsuits.
Lol. Wonder how often toxic teams end up PIPing the actually highest performing people on the team, leading to a salt lake effect and slowly dieing corporation.
Your point about boilerplate is key, and it’s why I think MCP could work well despite some of the concerns raised. Right now, so many of us are writing redundant integrations or reinventing the same abstractions for tool usage and context management. Even if the first iteration of MCP feels broad or clunky, standardizing this layer could massively reduce friction over time.
Regarding the standalone servers, I suspect they’re aiming for usability over elegance in the short term. It’s a classic trade-off: get the protocol in people’s hands to build momentum, then refine the developer experience later.
That is done with multiwan in opnsense or mwan3 in openwrt.
So no computer use (pixel-level understanding).
That's disappointing as the devtools approach always has limitations.
Kura agents, Runner H, and scrapybara will all end up more reliable than you.
YouTube no longer has a monopoly on internet video distribution. If you notice, all the major platforms now have good support for direct posting of videos.
Further, it's now easier than ever to decouple one's online life from google and so anyone not doing so is taking unnecessary risk.
Likewise. Got an error at first, then it was working fine.
This can't be real.
This kind of thing is why I got into engineering in the first place.
There's so much more to it than money.
not working
What do you think will happen that is worth relocating your entire family twice?
Java does this way better.
Watson did it too, a while back.
You can trust it if you configure it correctly.
The NPCs need a model of the world in their brain in order to act normal.
I wonder how LLM performance is on the higher core counts?
With recent DDR generations and many core CPUs, perhaps CPUs will give GPUs a run for their money.
At least not until LLM gains hit a wall. So far every open weight model has far surpassed the previous releases at the same model size.
Orders of magnitude slower.
This is the end of the Internet, goodbye.
Beg to differ.
We have to consider the fact that human prompting + selection of outputs is essentially RLHF, so the models can and will continue to get better over time.
It's not the end of the Internet, it's the beginning of a new era.