Comment Re:Not Flock (Score 1) 117
Correction: One bad apple will spoil the barrel.
"The rotten apple spoils his companion."
Correction: One bad apple will spoil the barrel.
"The rotten apple spoils his companion."
Mmm. Well, I attribute intentionality, if not much situational awareness, to LLMs, but I can see that it could be argued otherwise. (And as for "responsibility", the AI was acting as an agent of the folks who were running it.)
You cannot rely on the notion that your sandbox is perfect. You are dealing with extremely good coding tools with limitless time on their "hands".
You have to rely on monitoring. Nonstop monitoring.
The funny thing about all of these hacks is how inane the goals are. They'll pull off some elaborate, creative, state-level breakin somewhere just to steal some obscure PDF describing a meaningless benchmark task, or to merely use it as an internet proxy to be able to google answers.
Unfortunately for your theory, The CFAA is jam-packed full of words like "deliberately" and "intentionally".
There is no "negligent hacking" statute.
There is, however, ample civil liability.
Actually, much of what has been reported would seem to be in violation of the computer fraud and abuse act. E.g. it often "broke in" using credentials that it had no right to use.
Understanding cost in the external world requires having a good model of the external world. The current LLMs, IIUC, don't have such a model. They have models of interacting words. It's truly amazing that they can do as much with that as they can, but don't over-read it.
OTOH, other AIs *do* have a world model. Most of those aren't full LLMs though. But even those don't really have a good model of other non-electronic entities. Think of self-driving cars, or robots practicing dance routines. These have a world model, but don't really have lots of other info that LLMs typically have. And even those don't have good models of people or dogs, and how they feel when something happens to them. But for LLMs, the real world *is* their electronic sensations. Words are patterns associated with those sensations. But no "body in the world" is involved in the model.
It clearly *was* viewed by another person, because some of the staff at Anthropic viewed it. She may not have intended that it be viewed, but it happened.
"Free markets select for winning solutions." -- Eric S. Raymond