HN user

mike741

135 karma
Posts0
Comments133
View on HN
No posts found.

Urgent updates can be necessary every once in a while but should be recognized as technical failures on the part of the developers. Failure can be forgiven, but only so many times. The comments saying "what about X update that had this feature I need?" are missing the point entirely. Instead ask yourself about all of the updates you've made without even looking at the patch notes, because there are just too many updates and not enough time. Instead of blaming the producers for creating a blackbox relationship with the consumers, we blame the consumer and blindly tell them to "just update." That's what needs to change. It's a bit similar to opaque ToS issues.

The test is novel to the program, just not its programmer. So are we testing the program or are we actually testing its programmer? If we're testing the program, then the programmer's foreknowledge is irrelevant.

If it is incorporated, its definitely not effective. Clickbait dominates Youtube's recommendations and search despite consistently low thumb ratings. An easy example would be a procedurally generated channel such as this one:

https://www.youtube.com/@futureunity5129/videos

Sort by "Popular" and you'll see that their most watched videos have consistently low like/dislike ratios yet are still being actively recommended. If you use Youtube's search feature, these same channels and videos will come up long before the actually informative channels do.

You could argue it's been helping Youtube by wasting viewer's time and making them watch extra ads but its certainly not helping the viewers find what they're looking for. Even mass reporting the channels doesn't seem to stop them.

if its "not that much" and creates deception then why make site-wide changes that make it even less transparent? If a video had 4k dislikes and 46k likes then you knew that 4k accounts disliked it. There's little room for speculation there. Now you have to go to the comment section to get an idea of public opinion, where the theorized vocal minority have far more potential influence (because they can leave multiple comments, but only one thumbs down)

LLMs can pass novel theory of mind tests, which is what we're talking about.

Passing a ToM test is not what OP meant by having an "underlying theory of mind." OP's talking about the machine having an underlying mind (ie sentience, sapience, consciousness, etc), ToM tests are only testing output.

You said "Those tasks could be completed a [sic] traditional static program.", and no, they can't. You're incorrect.

They can, a static program as I described would indeed answer that one question correctly, resulting in a positive ToM score, without seeing any training data whatsoever. Did the programmer see it? Maybe, but the machine didn't and it would pass the test regardless.