HN user

beering

3,121 karma
Posts4
Comments516
View on HN

Really awful how the AI labs are skillmaxxing /s

Pelicans aside, we need to remember that benchmarks are the only good quantitative way we have of comparing models. If someone has complaints about “benchmaxxing”, please ask them to contribute a better benchmark! It is valuable work and very appreciated.

I think the third party only has an estimate for web visitors. Most users are probably on the mobile app?

I think parent’s point is that every false conjecture can cost a lot of time to be spent on futile affirmative proofs. So if we “clean up” a bunch of false conjectures, then more effort can be spent on interesting proofs of the others. (Probably a rather naive view of the value of conjectures but I’m just offering an alternative interpretation of the comment.)

You really do not want to live in a world where people around you are as dumb as hell. You don’t want your car mechanic, your restaurant chef, your local traffic engineer to be dumb as hell because that only makes your life worse.

Closer to you, you don’t want your customers, your coworkers, your friends, and your neighbors to be dumb as hell. Whatever you may gain by being the smartest one on your block is more than offset by the myriad ways in which dumb people make life worse.

GPT-5.6 12 days ago

Your names are not good because “Fast” is not a descriptor of model size and overlaps with fast/ultrafast inference. And “Plus” collides with the ChatGPT subscription plan. Point being, naming is hard.

ChatGPT Work 13 days ago

because every non-programmer hears “codex” and thinks that it’s for coding only - seems like a large hurdle to adoption. claude has been successful with cowork branding which makes sense.

TFA doesn’t actually state where the bit about shockwave therapy came from and it wasn’t the main point of the article. The concern was about being given useless therapies. The homeopathic analgesic is concerning, at least to me.

I.e. nothing this radiologist said was related to the LLM’s advice.

I agree. For many people, LLMs are the first time that computers do what they tell them to. Not what some big tech PM has decided is or isn’t possible.

At the same time, OP is in the right to reject contributions they don’t want. Nobody providing open-source software is under any obligations to take changes. Forking is still a viable option in 2026. And I don’t think we need an on-demand app store either because the trust issues will still exist for good reason. We can have highly produced software coexisting with LLM agents.

Claude Fable 5 1 month ago

Right now there are Anthropic engineers deployed in the NSA to help them use their cyber models. The NSA is part of the department of war.

Because by definition, sapience is something only humans have. Ergo, parrots are not sapient.

More meta, all of the threads on this page are just people playing games with definitions. Eg, “qualia is something I have as a human but machines don’t have it. Therefore, LLMs do not have qualia.”

The set of tokens is learned, more or less. So I don’t get what point you’re trying to make here. There’s not a human manually deciding what tokens make up the token dictionary.

DaVinci Resolve 21 2 months ago

Photographers have already delegated their art to pushing buttons on a machine. They are the most receptive to AI tools, but not representative of artists in general.

FTFA:

I supported Polymarket for years because I believed they represented crypto values. In fact, their platform is simply another bucketshop where if you bet too much, you're going to get cleaned.

Please, tell me more about these “crypto values”. Are they values like, “no regulations,” “rug pulls,” “funding ransomware”?