The study is "new", but the data they are using is old. It has references GPT-4 for god sake.
HN user
strangescript
I understand the struggle, but I feel like so many of these projects are going to ban AI contributions right before it starts getting undeniably better than humans.
The fundamental flaw in all these style arguments against core employment is people's personal desire to have a straight forward path for survival that doesn't involve a lot of strategic thought.
Its hard for highly intelligent people to understand that others simply DGAF. They just want their paycheck and they don't want to think about it deeply.
Thought leaders have existed since the dawn of human history. There were always scuffles at the top, but there were plenty of people that just went about their business, filling their roles and doing their jobs for the tribe.
Agents are getting good but professing they are surpassing you in domain and architectural knowledge with no special prompting is basically self reporting at this point. That could be your job wasn't that complicated or your personal knowledge wasn't that strong, either way, same result.
Don't get me wrong, I am sure we will get to all three of these pillars, probably by next year. I am not naive.
Hopefully we never do something silly like making a lead pushing machine that operates at high velocity, then mass produce it, what a terrible precedence that would set.
this is apples to oranges, the flash version version a full version
this thing is like 5x better than flash at fine grain detail
Anecdotal, of course, but the biggest change I ever made in my life was right before bed: take a screaming hot shower with dim lighting. I'd say 95% of the time, I get in bed and just pass out and have no real memory of time passing before falling asleep.
Same reason why two companies have the same idea, one goes viral and one doesn't. Public opinion matters even if its illogical at times.
Their concerns weren't completely off base, I think they just over estimated how much it would really matter in the grand scheme.
I will "suffer" through .004 of electricity if I can run it on my own computer
When you cannot win, you regulate the scoreboard. This is why Europe keeps getting lapped. They mistake control for competence. They cannot build platforms at scale, so they try to govern everyone else’s.
If your co-workers are getting promotions and raises and you are not, its a you problem. If someone else is getting credit for your work, its a you problem. Given your claims of impeccable work, we are only left to assume its a personality issue.
Its not to say its fair or right, but life is a popularity contest, whether we like it or not. More likeable people get more things, sometimes undeservingly so.
If you read the PR, the bad issues are in a few extensions, not the bot itself. The unencrypted oAuth token isn't really a big deal. It should be fixed but its a "if this box is compromised" type thing. Given the nature of clawdbot, you are probably throwing it on a random computer/vps you don't really care about (I hope) without access to anything critical.
Same with food. Plenty of food, just not where its needed.
I find gpt-oss 20b very benchmaxxed and as soon as a solution isn't clear it will hallucinate.
Not really an apples to apples comparison. You are comparing it to core technologies that millions of things sit on. There will always be money for that.
Its almost never correct to rip on a project from a distance. Only two things can happen, one, you are wrong, and the project succeeds. This is a personal catastrophe for your career at that company. Two, you are correct and the project fails. Its rare this will get you enough credibility to make the risk worth it. There are always others that will show up and dogpile as if they "knew" the entire time themselves. You need to be consistently correct about failure to get truly noticed, but then it asks a lot of questions. Why are you still working there? Why don't you have enough influence to prevent it in the first place? "I told you so" rarely accomplishes anything good.
This is cool, but once a week seems a little slow
yes, AI isn't penetrating those fields with high job losses at all
Cool project, but I never found the cloudflare DX desirable compared to self hosted alternatives. A plain old node server in a docker container was much easier to manage, use and is scalable. Cloudflare's system was just a hoop that you needed to jump through to get to the other nice to haves in their cloud.
this, provided you don't mind hopping around a lot, 5 20 dollar a month accounts will get you way more tokens typically, also good free models will show up from time to time on openrouter
If your outlook is 10 years then for sure, its valid. I am not sure how you come to that conclusion logically though. At the beginning of the year we had 0 code agents. Now we have dozens, some are basically free, (of various degrees of quality, sure).
The last 2-3 months of releases have been an unprecedented whirlwind. Code writing will be solved by the end of 2026. Architecture, maybe not, but formatting issues isn't architecture.
Do they consider code readability, formatting and variable naming as "errors" for the overall count. That seems dubious given where we are headed.
No one cares what a compiler or js minifier names its variables in its output.
Yes, if you don't believe we will get there ever, then this is totally valid complaint. You are also wrong about the future.
AI can make you a basic signal for whatever group you want with zero oversight now anyway. The days of trying to proxy anti-encryption laws so you can spy on your people are numbered.
"This post is to underline that Nothing is inevitable."
Inevitable and being a holdout are conceptually different and you can't expect society as a whole to care or respect your personal space with regards to it.
They listed smartphones as a requirement an example. That is great, have fun with your flip phone, but that isn't for most people.
Just because you don't find something desirable doesn't mean you deserve extra attention or a special space. It also doesn't you can call people catering to the wants of the masses as "grifters".
I am saying it doesn't matter. Driving in motor vehicles is dangerous, but we still do it. It would be safer to never drive anywhere. We have decided its an acceptable risk. Someone doesn't post one car accident on hacker news where they got hit by a semi and we all collectively say "oh man, never driving again".
Your assumption is you are "in control", as does everyone right before they have an accident.
you should probably avoid driving or riding in motor vehicles
I work 60+ hours a week with Claude Code CLI, always run dangerously skip, coding on multiple repos, on a mac. This has never happened. Nothing remotely close has ever happened. I have been using CC since research preview. I would love to know the series of prompts that lead to that moment.
My kids have book reports and stuff. Lately I can use AI to generate non-trivial questions about the books and use it to quiz them without me knowing anything about the books. Been super useful.