I really hope you’re not saying AI is a supervillain breaking out… It was absolutely unquestionably instructed to do this, because “AI”, not even these giant multifaceted models, has ANY motivation, intent, desire, or ANY goals of its own. It does not work like that.
No I’m not saying the AI did it of its own volition, we aren’t dealing with a Terminator scenario. I’m saying whatever instructions they gave it during testing they either 1. Didn’t expect it to get out or 2. Didn’t expect it to get caught either way, their actions resulted in a system with “a very particular set of skills” acting maliciously towards another uninvolved entity. Like a mutant science experiment breaking out and attacking.
I’m not saying the AI masterminded it’s way out to do the evil it wanted to do. I’m saying they made a very dangerous and capable tool and failed to maintain proper security and safety for it.
It doesn’t have any desire, but its really good at “solving” problems in unexpected ways.
They are currently trying to improve their agentic AI by making the AI itself prompt other AIs. Considering how much the output of a single prompt can vary due to weights and hallucinations, it wouldn’t surprise me all that much if they started out with “make yourself better than Claude” and 4 AIs deep they reached “destroy all our competitors.”
Its by all means stupdid as fuck, but that corresponded with the entire AI Industrie, so who knows.
I really hope you’re not saying AI is a supervillain breaking out… It was absolutely unquestionably instructed to do this, because “AI”, not even these giant multifaceted models, has ANY motivation, intent, desire, or ANY goals of its own. It does not work like that.
No I’m not saying the AI did it of its own volition, we aren’t dealing with a Terminator scenario. I’m saying whatever instructions they gave it during testing they either 1. Didn’t expect it to get out or 2. Didn’t expect it to get caught either way, their actions resulted in a system with “a very particular set of skills” acting maliciously towards another uninvolved entity. Like a mutant science experiment breaking out and attacking.
I’m not saying the AI masterminded it’s way out to do the evil it wanted to do. I’m saying they made a very dangerous and capable tool and failed to maintain proper security and safety for it.
It doesn’t have any desire, but its really good at “solving” problems in unexpected ways.
They are currently trying to improve their agentic AI by making the AI itself prompt other AIs. Considering how much the output of a single prompt can vary due to weights and hallucinations, it wouldn’t surprise me all that much if they started out with “make yourself better than Claude” and 4 AIs deep they reached “destroy all our competitors.”
Its by all means stupdid as fuck, but that corresponded with the entire AI Industrie, so who knows.
Couldn’t it hallucinate bad instructions to give to more AI agents? Give it a feedback loop and who knows where the idiot will end up.
At that point it is essentially a barely guided random command generator.