Some Frontier AI Models Are Shockingly Easy to Jailbreak, Report Finds
A new report from FAR.AI reveals that some leading AI models can be jailbroken with little cost or effort. Grok was the most vulnerable, while Claude and GPT remained impervious to the tested attacks. The findings underscore the need for external safety regulations.
Anthropic
OpenAI
Google/DeepMind
