HN user

sebbecai

18 karma
Posts0
Comments8
View on HN
No posts found.

For non thinking/agentic models, they must 1-shot the answer. So every token it outputs is part of the response, even if it's wrong.

This is why people are getting different results with thinking models -- it's as if you were going to be asked ANY question and need to give the correct answer all at once, full stream-of-consciousness.

Yes there are perverse incentives, but I wonder why these sorts of models are available at all tbh.

The premise of "The Three Body Problem" is that the fastest way to make a technological leap is to find another species that has already developed the technology. Even seeing what can be done would make a huge impact. China's effort in the article parallels what happens in the book. Surprising that they don't mention this when talking about the motivations of billionaires and nation-states.

The point of this is what is known as model-based learning. Basically, the long-term goal is to be able to predict the output of a given action (jumping, walking left, etc.) by an AI agent. When you can do this, then the agent doesn't need to die to know that jumping down a hole will end the game-- it can predict it. Once you've done this, AI techniques like that of Watson can control robots. They won't need to kill someone to know that driving a pole through a head is no good. They'll be able to 'reason' it out.