Anthropic

Latest AI news, models and releases from Anthropic. ['anthropic-ai', 'Anthropic AI Models', 'Claude', 'Claude 2', 'Claude 2.1', 'Claude 3', 'Claude 3.5 Sonnet', 'Claude-3.5-Sonnet', 'Claude 3.5 Sonnet v2', 'Claude 3.7 Sonnet', 'Claude 3 Haiku', 'Claude 3 Opus', 'Claude3 Opus', 'Claude 3 Sonnet', 'Claude 4', 'Claude 4.6', 'Claude 4 Opus', 'Claude 4 Sonnet', 'Claude 5', 'Claude Agent SDK', 'Claude.ai', 'Claude AI', 'ClaudeBot', 'Claude Code', 'Claude Code (Opus 4.7)', 'Claude Code Sonnet 5', 'Claude Cowork', 'Claude Design', 'Claude Desktop', 'Claude Fable', 'Claude Fable 5', 'Claude Fable5', 'Claude Fable 5 Max', 'Claude Haiku', 'Claude Haiku 4.5', 'Claude Instant', 'Claude (language model)', 'Claude Max', 'Claude Mythos', 'Claude Mythos 5']

Research 🇺🇸

Society can be reward-hacked, just like cyber environments:…Imagine an army of credit card point optimizers gaming the system… forever…

Researchers from Kings College London, Fudan University, and The Alan Turing Institute developed SocioHack, a benchmark testing AI's ability to 'game the system' in real-world scenarios like credit card points and school grades. Meanwhile, Anthropic reported an 8x increase in code merged in 2026 vs 2021-2024, suggesting prosaic recursive self-improvement. Additionally, University of Zurich and Google DeepMind demonstrated RL-trained drones that outperformed a human champion drone racer, with implications for warfare.

AnthropicAnthropic Google DeepMindGoogle DeepMind
Import AI27.07 · 12:03
Fresh news