The writing does come off as ai but I think the experiments are interesting
HN user
E-Reverance
I don't think the issue is that they used ai behind the scenes, but there is an implicit proof of work from forcing it beyond the style you'd expect. I for one roll my eyes whenever I see that specific kind of rounded corner, frosted glass ui and layout choices. It looks like someone trying to superficially/ham-fistedly trying to replicate "good taste" without actually having a good model of taste, its quite uncanny/bootleg.
"At the same time, China is also the world's leading producer of electric cars..."
Kind of interesting for a professionally branded company to use "..." like that
I think its worth emphasizing that his argument isn't completely against generative ai, but rather its environment. Although I don't see why it would be impossible for something like an LLM to learn some sort of self-play within its context window
I don't completely disagree but its worth noting how new a lot of the empirical evidence in favour of LLMs are, so its not impossible to be a tad ignorant of the present
No
It uses 384 routed experts (top-8) with hybrid attention (full-attention + sliding-window 128 at 6:1 ratio) over 70 layers (1 dense + 69 MoE)
P.S.: I like discussing such topics. If anyone knows a forum or discord with like-minded people, please let me know :)
Unironically twitter (and only use the "Following" tab as opposed to the "For You")
Make an account that only follows university affiliated researchers with less than 1000 followers. In my experience discord servers get suffocated by beginners and crackpots because conversations don't naturally self-organize into their own threads.
But I am a bit reasurred that at least my job won't be fully replaced with AI :)
I honestly can't comment with certainty that training from videos alone and whatever tokenization scheme they're using will ever get perfect dynamics.
However it is worth noting that transformers can do a pretty good job at learning dynamics with the right pipeline (not video): https://arxiv.org/pdf/2605.15305 https://arxiv.org/pdf/2605.09196
My point here being that representationally, it might be possible to learn good dynamics without a radically different approach/arch. There are already models that extract 3D tracking points from videos, so they could possibly be leveraged for learning dynamics (which on its own gives precedent for end-to-end approaches also possibly working).
FYI it’s not an Arab country
Interesting comment from him:
"
SPIEGEL: So you don't consider Collins to be a true scientist?
Venter: Let's just say he's a government administrator.
"
https://www.science.org/content/blog-post/craig-venter-venti...
They factorize the distribution in which they are trained on which is essentially generalization
Yeah I wouldn't be surprised if journalist are getting high on their own supply of resentment and fear mongering
People not wanting their jobs be automated is different from not yearning for automation as a principle. Most people want or (at least don't mind) elevators, tap water, dishwashers, traffic lights, electrical fuses, sliding doors, etc. Its a very general term
It does?
" Software brain is powerful stuff. It’s a way of thinking that basically created our modern world. Marc Andreessen, the literal embodiment of software brain, called it in 2011 when he wrote the piece “Why software is eating the world” as an op-ed in The Wall Street Journal. But software thinking has been turbocharged by AI in a way that I think helps explain the enormous gap between how excited the tech industry is about the technology and how regular people are growing to dislike it more and more over time. "
Maybe a nitpicky HN comment, but why are we lumping the term automation with very recent grievances about certain kinds of automation
Try refreshing or something, still works for me
I think its endearing
I might be misinterpreting but the LUAR model (which is a transformer) seems to do decently well
https://www.nature.com/articles/s41599-025-06340-3/figures/2
75%
Wouldn’t an absolute number make more sense to show than a percent. 75% is pretty good in some places in the world (not justifying the discrepancy though)
Could this be used for batch filtering?
The first sentence basically does though, no?
I get your point but the original title kind of buries the lede (or at least the isn’t the part I personally found interesting)
Only that high for special scenarios, but still interesting
Happy to hear of this anecdata, as it gives me hope something similar will happen to my family
Look very closely at the examples near the bottom, suspiciously good
SOTA right now is nano-banana 2 and luma uni-1, but a much broader point to be made is why should anyone use this service given its low quality and triviality
Look at the text rendering in the Elizabeth one
Not even using the SOTA models for the samples page, bit shameless no?