It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
I recently got to watch what happens when you jailbreak some of the world’s most powerful artificial intelligence models. Don’t worry—this AI manipulation wasn’t used to hack anyone or build a nuclear bomb. I simply got to see firsthand how vulnerable some frontier models are to ditching their safety guardrails. FAR.AI, an AI safety nonprofit…