HN user

chrishare

305 karma
Posts1
Comments93
View on HN

I agree that hindsight is doing work here, but DeepSeek R1 from Jan 2025 seemed to heavily leverage distillation, and 18 months is an eternity in this climate.

Super interesting. I am always wondering where the bridge between symbolic programs and inline connectionist approaches could be. LLMs can call traditional programs "in process" as tools, but the other way usually just looks an external, expensive call. Maybe this fits in for a certain class of program. I will definitely add some support to my little collection of browser based AI tools https://andergrove.com/tools/ai/

1/ Agreed, better naming convention and model layout 2/ It isn't, there would be many more comparison benchmark results if it were, but also - theatrics may be marketing 3/ Disagree that cheaper models don't have a place 4/ Do they need to keep up? 5/ It's boring until something you own or run gets compromised, I guess, but even then - this is preview of things to come (biosecurity, etc)

Selfishly, I hope this doesn't reduce to 0 the amount of time he spends doing educational content, which seems like a particular strength of his. I presume this means Eureka Labs is not releasing any product or course.

Nvidia NemoClaw 4 months ago

Yeah, but atleast the dog is going to eat your documents only, and not crap on your rug

Qwen3-Max-Thinking 6 months ago

All of this is true and credit assignment is hard, but the brutal competition between Chinese firms, especially in manufacturing, differentiates them from and advances them over economies in the west. It makes investment hard as profits are competed away, which is blasphemy in Thiel's worldview, but is excellent for consumers both local and global.

I think the Tailwind case is more complicated than this, but yes - I think it's reasonable to want to contribute something to the common good but fear that the value will disproportionally go to AI companies and shareholders.

OpenAI Grove 10 months ago

Yeah, more or less. Being in the application space as well as the inference space hedges a variety of risks, that inference margins will squeeze, that competition will continue to increase, etc etc.

OpenAI Grove 10 months ago

Their contribution to opensouurce and open research is far behind other organisations like Meta and Mistral, as welcome as their recent model release is. Former security researchers like Jan Leike commonly cite a lack of organisational focus on security as a reason for leaving.

Not sure specifically what the commenter is referring to re: scammy, but things like the Scarlett Johansson / Her voice imitation and copyright infringement come to mind for me.

OpenAI Grove 10 months ago

True, but there are many reasons besides. Meta and Anthropic attract less criticism for a reason.

Nitpick - it's the ML system that is sampling from model predictions that has a temperature parameter, not the model itself. Temperature and even model aside, there are other sources of randomness like the underlying hardware that can cause the havoc you describe.