HN user

dkdcwashere

46 karma
Posts0
Comments10
View on HN
No posts found.
AI 2027 1 year ago

The alignment community now starts another research agenda, to interrogate AIs about AI-safety-related topics. For example, they literally ask the models “so, are you aligned? If we made bigger versions of you, would they kill us? Why or why not?” (In Diplomacy, you can actually collect data on the analogue of this question, i.e. “will you betray me?” Alas, the models often lie about that. But it’s Diplomacy, they are literally trained to lie, so no one cares.)

…yeah?

the author ironically uses a lot of words to say very little, though I agree with the conclusion. it’s already annoying to have someone use a lot of words to say very little (especially in a business context). now it’s free and easily accessible for anyone, whereas before it at least took some social stamina

so people will do it, people will be annoyed by it, people will prioritize to more efficient communicators

GPT-4.5 1 year ago

*good*. the answer to this is legislation —- legally, stop allowing shitty ads everywhere all the time. I hope these problems we already have are exacerbated by the ease of generating content with LLMs and people actually have to think for themselves again