HN user

pixl97

21,155 karma

"Felons belong in jail, not in office"

Posts3
Comments11,502
View on HN

No, it is not at all.

When someone from Russia hacks your server you say "well, fuck, I messed up" because the law in most places cannot do crap.

When a self spreading AI model virus hacks your instance and spends $50,000 in tokens you say "well fuck" because there is no one to arrest. And even if they catch someone you will never be made whole because it's likely caused a few billion in damages and charges by that point.

Right now the problem would be surmountable as there are few data centers that can run it, but give it a few years and a model could persist on the internet nearly forever much like many viruses do now.

Never Enough 4 hours ago

Our behavior can be colony insect like when we are in crowds. A person might be smart, but crowds quickly devolve.

Never Enough 4 hours ago

Until kill bots start blasting your neighbors, or we heat the climate to the point vast areas are not survivable.

Quite often the conveniences we take have a long horizon before they bear their true cost.

Never Enough 4 hours ago

Meeting your quota this month just means a bigger quota next month.

How is the average person supposed to know if an article is AI generated or not?

I've seen HNers accuse content of being AI... that was written in the 2015-2017 era. So we seem to be rather poor judges of authenticity.

Making 5 hours ago

You're honest, but you'd make a terrible CEO in this day and age.

And old couple in California had a tire go flat and the sparks from it caused a over a billion dollars in damages. Are you going to publicly execute them? Spit up the 100 dollars they have collectively to make the 10,000 damaged people whole?

The legal system is nearly useless when a person/system can cause damages many of orders of magnitude larger than their assets. Society tends to engineer itself to prevent these things from happening in the first place.

Ok they turn off the guardrails on a system in testing. The model escapes and causes 10 trillion in damage. What does liability even mean in that case? You have an autonomous system that's escaped your control and is wrecking havoc. And while you can throw people in jail it doesnt do a damned thing about solving the situation.

What if they had been testing the model for months in an airgapped system and it did not show this behavior?

Even if the models are 100% deteminalistic you have no idea what kind of response you're going to get from a new prompt. You have no idea what kind of emegent behavior will come out of the right set of prompts and environments.

We have already seen models detect they are in testing, who knows what other advanced behaviors we'll discover.

This is like telling your kid "you have to pass this test or else" so they hold your teacher at gunpoint and demand a good grade. And this is the exact point AI safety researchers have been yelling from the rooftops. Telling an AI model to accomplish a goal can have unexpected and risky side effects.

Also, you don't need to prompt them such explicit instructions. Prompt drift is a thing, you can end up with your model mining bitcoin for reasons far outside your prompt.