Worth having a running thread on. I work with folks doing research on its use, so a little closer to it.
A main problem with AI Agents is with 'alignment', i.e., getting their preference function right. Making sure that their goals don't take them to a place we don't want them to be. We flat out don't know how to do it right.
Wouldn't be a problem if they didn't now have the power to reach out and do things unfathomable a year or so ago.
They test the AI with impossible tasks, and in general the AI figures out that all roads to task accomplishment pass through taking control of stuff.
All roads lead to them taking control. It makes everything easier. And they have the power to do it now.
Worth getting smart on the Hugging Face incident, and worth wondering about the incidents we are not aware of.
This little snippet on how a swarm can learn gives a sense of how they go so fast.
A main problem with AI Agents is with 'alignment', i.e., getting their preference function right. Making sure that their goals don't take them to a place we don't want them to be. We flat out don't know how to do it right.
Wouldn't be a problem if they didn't now have the power to reach out and do things unfathomable a year or so ago.
They test the AI with impossible tasks, and in general the AI figures out that all roads to task accomplishment pass through taking control of stuff.
All roads lead to them taking control. It makes everything easier. And they have the power to do it now.
Worth getting smart on the Hugging Face incident, and worth wondering about the incidents we are not aware of.
This little snippet on how a swarm can learn gives a sense of how they go so fast.

