How long until Gemma 5 hits?
HN user
ishurand4
3.1 Flash Lite has a better pelican that 3.6?
The only one I see that thinks it is claude other than claude itself is the GLM series.
Its a quadratic graph. It starts low but not that capable, gets better and more expensive, and then the time comes in which the capability needed is not the ones of the frontier models and then the price goes down on the companies who host the models that the capability is "good enough"
A myth, not a model that actually exists
A system "card" made mostly by the model itself.
Thats why it is a mythos model
Well, for me at least, I pay more for input (Up to 1M per prompt) than output (usually max 4k-8k)
Yeah, it is also known as Claude Mythos 5
What's the token usage at?
The numbers they show don't matter. "On multi-round coreference/context recall tests (often cited as MRCR or long-text retrieval benchmarks), Opus 4.7 reportedly dropped from roughly 78.3% down to 32.2% compared to Opus 4.6.", but what did anthropic do? They just stopped showing the benchmark altogether and then just show the cherry top ones that got improved on.
And anyway, with quantum, there will be no need for frontier companies as you might be able to even run a 1T param model on a consumer quantum computer.
They just showed the benchmarks it improved on but it regressed on so much more, such as the MCRR benchmark: "On multi-round coreference/context recall tests (often cited as MRCR or long-text retrieval benchmarks), Opus 4.7 reportedly dropped from roughly 78.3% down to 32.2% compared to Opus 4.6."
Stack Overflow and Reddit are still getting threads. And as AI gets smarter, the questions will also expand.
But then they have alot more services since msft bought them, making higher chances on goes down