HN user

reckless

67 karma
Posts0
Comments25
View on HN
No posts found.
Qwen 3.8 4 days ago

5.6-sol would be a better comparison given it's general availability and usage allowances

Takeaways Which model to use for generating architecture.

The simpler the decision you need to make, the more readily you can just take a proposal from Fable or GPT-5.5. And now that Fable is unavailable outside the US, the choice between Opus and GPT is far from obvious. As a quick default I'd lean toward GPT.

Claude Opus 4.8 2 months ago

No way is Muse Spark generally better than offerings from Google and OpenAI. I actually find arena to be amongst the most useless indicators

The aggregate picture only tells you so much.

Sites like simonwillison.net/2025/jul/ and channels like https://www.youtube.com/@aiexplained-official also cover new model releases pretty quickly for some "out of the box thinking/reasoning" evaluations.

For me and my usage I can really only tell if I start using the new model for tasks I actually use them for.

My personal benchmark andrew.ginns.uk/merbench has full code and data on GitHub if you want a staring point!

True but what you can do is SSH to the device and install a custom launcher for apps that can read standard epubs, play chess, or expose the linux terminal on device.

Not great for basic users but I've had significantly more use out of it with some advanced setup.

Seems to be entirely a different approach for diffusion.

DeepFloyd IF works in pixel space. The diffusion is implemented on a pixel level, unlike latent diffusion models (like Stable Diffusion), where latent representations are used.

I feel like the comments on backwards compatibility are due to the absolute shitshow of TF2 compatibility for TF1 code and models.

Also the threat of Pytorch can be seen when reading between the lines, especially since it's now run by a foundation and the darling of the diffusion model developments.

TL;DR is that AVX-512 provides >20% speedup compared to AVX2

It's a shame that intel isn't including AVX-512 on the 13th gen Raptor Lake consumer CPUs, and the newer skews of 12th gen like the KS processors have it fused off.