Regardless of the marketing material, LLMs are still just using probabilities to guess what the next token is.
HN user
ru552
You won't like it, but the answer is Apple. The reason is the unified memory. The GPU can access all 32gb, 64gb, 128gb, 256gb, etc. of RAM.
An easy way (napkin math) to know if you can run a model based on it's parameter size is to consider the parameter size as GB that need to fit in GPU RAM. 35B model needs atleast 35gb of GPU RAM. This is a very simplified way of looking at it and YES, someone is going to say you can offload to CPU, but no one wants to wait 5 seconds for 1 token.
Gemma 4 has made a lot of progress in this area. The model is phenomenal. It's size is workable. This is the worst it will ever be.
There's speculation that next Tuesday will be a big day for OpenAI and possibly GPT 6. Anthropic showed their hand today.
The article says the policy change is separate and unrelated to Anthropic’s discussions with the Pentagon.
the article specifically says:
The policy change is separate and unrelated to Anthropic’s discussions with the Pentagon, according to a source familiar with the matter.
Virustotal at upload and periodically during the day
sure it does, Bezo's space company and Google are both planning the same
Here's Sundar talking about doing it by 2027: https://www.businessinsider.com/google-project-suncatcher-su...
You can talk to it in discord or whatsap or telegram etc. cause it's checking for you in a loop.
That's the biggest difference I can tell.
It was tough, but it wasn't Battletoads tough.
I like this workflow
"The constraint system offered by Guidance is extremely powerful. It can ensure that the output conforms to any context free grammar (so long as the backend LLM has full support for Guidance). More on this below." --from https://github.com/guidance-ai/guidance/
I didn't find any more on that comment below. Is there a list of supported LLMs?
*Please note, I'm not in favor of censorship, it's just that this analogy is inaccurate
Olive Garden isn't given access to something it requires to operate at the pleasure of the government. Broadcast TV on the other hand...
All of broadcast TV is allowed because the government says it is. ABC/CBS/NBC/FOX don't own the radio spectrum they are operating on, the government does and they grant the right to use it to those companies. There's a long list of things that the government requires them to do in order to keep this pleasure. One of them used to be the Saturday morning cartoons. I miss those.
Generally between 11a and 7p. Going to lunch/dinner.
I've used Waymo countless times in SF. It's typically 15% cheaper than an Uber/Lyft and trip time/wait are generally the same. I much prefer the Waymo.
You're better served using Apple's MLX if you want to run models locally.
This is the model that was code named "Sonic" in Cursor last week. It received tons of praise. Then Cursor revealed it was a model from xAI. Then everyone hated it. :/ I miss the days where we just liked technology for advancement's sake.
*edit Case in point, downvotes in less than 30 seconds
This explains why Anthropic cut OpenAI off of their models just before GPT-5 was released.
Is this much different from the custom GPTs that OpenAI pushed a year or two ago?
say what?
I wonder if this meets the requirements set by the recent round of outside investors?
Sure, but you can't use it until the hardware shows up.
I think FAL is already too big for a "viber" company to buy. $55M ARR and 10%+ EBITDA. They're either going to IPO or get acquihired by a big 4.
How are they giving AWS a run for their money when they use AWS for their own service? AWS profits from Supabase growth.
My humble guess is they sold themselves as the enabler for the AI vibe coding "revolution"
They only stopped referring to debt collections in 2020 when Covid hit.
Considering that they are the model that powers a majority of Cursor/Windsurf usage and their play with MCP, I think they just have to figure out the UX and they'll be fine.
Copy in copyright is not copy like copy in copying some data.
Copy in copyright is a term for the actual writing that gets published on ads, or magazines, or in a news paper. "I need to get the copy from marketing for this campaign." "The editor hasn't approved the copy for the article yet."
Typically, people not in/around the industry aren't familiar with the term, which leads to the confusion.
that's where I stopped reading. that let me know this was not a rigorous study and had other motivations
It's been confirmed to run on a machine with no internet access. So it isn't reliant on external requests, though it could still be trying to make them.