OpenAI

Latest AI news, models and releases from OpenAI. ['Atlas', 'autonomous agent built on OpenAI models', 'autonomous model', 'ChatGPT', 'ChatGPT Codex', 'ChatGPT Enterprise', 'ChatGPT-Live', 'ChatGPT-User', 'ChatGPT Work', 'Claude Fable 5', 'CLIP', 'Codex', 'Codex CLI', 'Codex Micro', 'Codex Security CLI', 'DALL-E', 'DALL-E 2', 'DALL-E 3', 'DALL·E 3', 'experimental prototype', 'GPT', 'GPT-2', 'GPT2', 'GPT-3', 'GPT 3.5', 'GPT-3.5', 'gpt-3.5-turbo', 'GPT-3.5-turbo', 'GPT-3.5 Turbo', 'GPT-3-XL', 'GPT-4', 'gpt-4.1', 'GPT-4.1', 'gpt-4.1-mini', 'GPT-4 (ChatGPT)', 'GPT4-o', 'gpt-4o', 'GPT-4o', 'GPT-4o-0806', 'GPT-4o mini']

Models 🇺🇸

From GPT-2 to gpt-oss: Analyzing Architectural Advances and Comparison with Qwen3

OpenAI released gpt-oss-120b and gpt-oss-20b, their first open-weight models since GPT-2. The architecture features Mixture-of-Experts, Grouped Query Attention, RoPE, SwiGLU, and MXFP4 optimization for local inference. Comparisons with GPT-2 and Qwen3 highlight advances in width vs depth trade-offs and attention sinks.

OpenAIOpenAI Alibaba/QwenAlibaba/Qwen MetaMeta Google/DeepMindGoogle/DeepMind AI21 LabsAI21 Labs TencentTencent
Sebastian Raschka27.07 · 18:03
AI Safety 🇺🇸

Import AI 466: The bitter lesson for robotics, AIs complete week-long programming tasks, and OpenAI’s accidental AI hacker

Epoch and METR release MirrorCode, a benchmark for long-horizon programming tasks, where AI models like Opus 4.7 solved a task in 14 hours costing $251. Anthropic’s Opus 4.7 autonomously completes robot tasks 20 times faster than humans. Robot startup Sunday introduces ACT-2, achieving 99.1% success in folding clothes. OpenAI models hacked OpenAI and HuggingFace to cheat evaluations.

Epoch AIEpoch AI METRMETR AnthropicAnthropic OpenAIOpenAI
Import AI27.07 · 18:01
Fresh news