given this, is it safe to assume that inference pricing is barely related to cost to serve at this point and there is considerable margin?
HN user
anthonypasq96
the whole point of arc-agi 3 is that if models are AGI then they should be able to solve the same tasks as humans do given the same information, but they cant. allowing scripts and harnesses and whatnot completely defeats the purpose.
im a java guy, but it simply takes too long to start up and uses too much memory to be a reasonable language for ai agent development.
is your job to write code or develop software that works?
have fun keeping a job doing 1/3 the work of people getting paid the same as you :)
why does every AI skeptic assume that everyone is lying to them. theres millions of developers using AI to be more productive and you just keep plugging your ears and screaming, claiming its only dumb managers, meanwhile Linus Torvalds is vibe coding stuff.
knowing how to properly use AI is a skill, there are new tools, new patterns, new primitives etc. you will be unpracticed.
i agree, and its strange that this failure mode continually gets lumped onto AI. The whole point of longer term software engineering was to make it so that the context within a particular persons head should not impact the ability of a new employee to contribute to a codebase. turns out everything we do to make sure that is the case for a human also works for an agent.
As far as i can tell, the only reason AI agents currently fail is because they dont have access to the undocumented context inside of peoples heads and if we can just properly put that in text somehwere there will be no problems.
yeah, thats why people just use Teams
This is entirely too charitable. Basically all this proves is that the agent could run in a loop for a week or so, did anyone doubt that?
yes, every AI skeptic publicly doubted that right up until they started doing it.
i know you could do it, im asking why on earth you would feel its vital to verify stream.filter() was called twice in a function
what happened to not testing implementation details?
so 99% of all software?
youre verifying std lib function call counts in unit tests? lmao.
you sound insecure. that guy was making a thoughtful self-reflective observation and it seems like he hit a nerve.
i was a recipient of dreadful christrian brainwashing for about 14 years of schooling, and i got a fairly mild case
why are people online obsessed with the idea that anyone who disagrees with them is a paid actor
brother, no one cares. if LLMs made something exist that did not exist previously, they worked. it doesnt matter if you could have done it faster by hand if doing so would have resulted in the program not existing.