It depends on what you need, but Krea/Klein9b/Ideogram4/Z-Image are among the best right now for text2image and Qwen Edit and Klein are probably still the best at editing.
HN user
kroaton
Yup. Smells like marketing.
They've been investing heavily into this over the last 2 releases, but it's just that SideFX is a really small company. I think they have 50 people or so right now. Karma is way more accurate than Cycles, but slow. Karma XPU is a mixed bag.
And that's when you learn Houdini and open up a workflow from 2008 and it still works the same, despite SideFX pushing out an insane amount of features every release (look at their SneakPeaks).
Work more for my boss and landlord so the fiefdom can survive.
It also goes to show that Fable/Sol must be 4-5T in size.
Ling/Ring 1T-A50B and the new Inkling 975B-A41B deserve to be on that list.
Which would be very interesting to test, as larger models (such as Deepseek V4 Flash or Qwen 397B) seem to compress better. Their Q2 quants are usable as is, even without the ternary compression.
A shame it has Marc Maron in it.
Please don't use that garbage. Just use the base Qwen models or Nex/Orinth, as those are the only properly post-trained finetunes. The Qwopus models are marketing.
"DeepSeek-V4-Flash will fit" At Q2, 2bit? Lobotomized to death.
NVFP4 will be better if the model provider actually post-trained properly after quantizing.
You're late to the party, mate; we've been doing this for years. Grab a SearXNG instance, stand up an MCP server for it, and expose the tool into your system prompt. Or use Brave Search. Or Exa if you want to pay. Any of them work. The model will pick it up straight away.
Even llama.cpp's bundled web UI handles it fine. Dead simple.
This has been answered many times already. The short version: Brave's adblocker is not an extension, so manifest v3 has no effect on it.
To be fair, GPT5.5-Xhigh is similarly capable and has not burned the world down.
Anthropic nuked a big chunk of that "developer sentiment" when they rug-pulled us with the rate limits and gaslit us with "it was just a bug, guys!".
How is this not getting any traction? This is a massive problem. Pretty much every Reddit news app out there is now dead.
Did you even use it? It was nerfed to hell and back. It stopped following instructions, forgot what sub-agents responded and so on. Stop spreading this pro-Anthropic narrative. They did a rug pull due to lack of compute.
If Tauri ever gets proper webgpu support, that'll be the Electron killer.
Claude Code poisons non-anthropic models in usage. We found this out when the code was leaked. Use a fork or OpenCode/pi-coding-agent
Ask western models about Israel's genocides and mass rapes in Palestine, Lebanon, etc.
Buy any Strix Halo box and have fun with your 128GB of VRAM.
A3B-35B is better suited for laptops with enough VRAM/RAM. This dense model however will be bandwidth limited on most cards.
The 5090RTX mobile sits at 896GB/s, as opposed to the 1.8TB/s of the 5090 desktop and most mobile chips have way smaller bandwith than that, so speeds won't be incredible across the board like with Desktop computers.
For autocomplete, Qwen 3.5 9B should be enough even at Q4_k_m. The upcoming coding/math Omnicoder-2 finetune might be useful (should be released in a few days).
Either that or just load up Qwen3.5-35B-A3B-Q4_K_S I'm serving it at about 40-50t/s on a 4070RTX Super 12GB + 64GB of RAM. The weights are 20.7GB + KV Cache (which should be lowered soon with the upcoming addition of TurboQuant).
I did the same a few months ago when I read that multiple big OSS Linux projects were moving to it and it's been phenomenal so far.
It could just as easily be a $3000-4000 Strix Halo laptop.
If SPTM is active on the chip, we are not going to be getting Linux at all.
Loving this; great work! Do you talk about the process anywhere in more depth?
I remember thinking the same thing and this articles goes over most of the arguments here - https://milvus.io/blog/why-im-against-claude-codes-grep-only...
When it came out and I think it was Boris from Anthropic that said they experimented a lot with Vector Search and grep just worked better.
You can try it out using the Claude-Context MCP - https://github.com/zilliztech/claude-context