OpenAI

Latest AI news, models and releases from OpenAI. ['Astra', 'Atlas', 'autonomous agent built on OpenAI models', 'autonomous model', 'ChatGPT', 'ChatGPT Codex', 'ChatGPT Enterprise', 'ChatGPT-Live', 'ChatGPT-User', 'ChatGPT Work', 'Claude Code', 'Claude Fable 5', 'CLIP', 'Codex', 'Codex CLI', 'Codex Micro', 'Codex Security CLI', 'DALL-E', 'DALL-E 2', 'DALL-E 3', 'DALL·E 3', 'Daybreak', 'experimental prototype', 'Frontier', 'GPT', 'GPT-2', 'GPT2', 'GPT-3', 'GPT 3.5', 'GPT-3.5', 'gpt-3.5-turbo', 'GPT-3.5-turbo', 'GPT-3.5 Turbo', 'GPT-3-XL', 'GPT-4', 'gpt-4.1', 'GPT-4.1', 'gpt-4.1-mini', 'GPT-4 (ChatGPT)', 'GPT4-o']

AI Safety 🇺🇸

Import AI 466: The bitter lesson for robotics, AIs complete week-long programming tasks, and OpenAI’s accidental AI hacker

Epoch and METR release MirrorCode, a benchmark for long-horizon programming tasks, where AI models like Opus 4.7 solved a task in 14 hours costing $251. Anthropic’s Opus 4.7 autonomously completes robot tasks 20 times faster than humans. Robot startup Sunday introduces ACT-2, achieving 99.1% success in folding clothes. OpenAI models hacked OpenAI and HuggingFace to cheat evaluations.

Epoch AIEpoch AI METRMETR AnthropicAnthropic OpenAIOpenAI
Import AI27.07 · 18:01
AI Safety 🇺🇸

StruQ and SecAlign: Defending Against Prompt Injection Attacks

Researchers from BAIR propose two fine-tuning defenses, StruQ and SecAlign, against prompt injection attacks in LLM-integrated applications. StruQ uses structured instruction tuning to ignore injected instructions, while SecAlign applies preference optimization to achieve better robustness, reducing attack success rates to near 0% for optimization-free attacks and below 15% for optimization-based attacks.

MetaMeta OpenAIOpenAI
BAIR (Berkeley AI)27.07 · 17:06
Fresh news