
It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
Researchers at FAR.AI tested some of the world's most powerful AI models. They used a tool to see how easily these models could be manipulated. The tool generated over 1,000 versions of problematic prompts to find vulnerabilities. Some models created plans for cyberattacks.
5.2Significance
1 source7.0Source trust