

I don’t know about Gemini because I haven’t used it, but the other frontier models are absolutely smart enough to do some damage. If your only exposure to them is playing with the chat feature and asking it how many rs are in strawberry, you’re getting a skewed impression, IMO, and I’m not surprised you think they’re dumb.
The best models are getting scary good at programming. Not to the point of entirely replacing humans, but to the point where I don’t think software engineering will ever go back to being done by “hand.”
I think it comes down to the fact that programming tasks are inherently testable and verifiable in a way that most other tasks simply aren’t. The AI can write and run tests that directly tell it if it needs to adjust its approach. You don’t have that with prose or other less quantifiable tasks.
Sorry, I’m rambling a bit here, but my point is that hacking is wicked close to programming in skill set and in that you get really quick feedback if your approach is good or not. So, when you have an AI system that can just keep trying things essentially indefinitely, it’s scary good at compromising other systems and can do way more than the funny demos of LLMs confidently asserting very obviously wrong facts will make you believe.
Oh completely agree. My original post was implying Google allowed this to happen on purpose because they wanted the headline. Sorry if that wasn’t clear up front.