different system prompts, codex will have system prompt telling the model to gather a lot more context before starting work
HN user
rootatixww3
for that you would need to compare the same task implemented in two different languages - C# and Python for example, no?
they explain this is a benchmark, all models/harnesses receive the same prompt
chudo missed opportunity
don't forget "where are all these beautiful apps that supposedly everybody vibe codes now?"
yes, defense in depth
easier to plan/estimate compute for 5 days than for a month. worst case you only have 5 "unprofitable" days
there are many sources saying that Anthropic has 80% margin (profit) on API tokens
and "accidentally" they forgot to disable it when releasing
good approach, but your security should not depend on your router anyway, you should be immune to attacks from it
an LLM can't access its high dimensional vectors any more than we can access whatever the brain is doing at a low level
all kind of math structures were found in mammals brains - fourier transforms (well, not exactly), ballistic equations, Gabor filters
who knows how exactly we approximate the magnitude of a math operations, maybe we also use helices
my point is that we dont know if what we discover the neural networks doing (helical manifolds) is actually the same thing a brain converges on, or not
and there is an implicit bias here - evolution created language, and we forced neural networks to also evolve to be good at it. so it wouldn't be surprising to find some convergence, this particular kind of language turned out to work well (words, linear sentences, grammar)