OpenAI

Latest AI news, models and releases from OpenAI. ['Astra', 'Atlas', 'autonomous agent built on OpenAI models', 'autonomous model', 'ChatGPT', 'ChatGPT 5.6', 'ChatGPT Codex', 'ChatGPT Enterprise', 'ChatGPT-Live', 'ChatGPT-User', 'ChatGPT Work', 'Claude Code', 'Claude Fable 5', 'CLIP', 'Codex', 'Codex CLI', 'Codex Micro', 'Codex Security CLI', 'DALL-E', 'DALL-E 2', 'DALL-E 3', 'DALL·E 3', 'Daybreak', 'experimental prototype', 'Frontier', 'GPT', 'GPT-2', 'GPT2', 'GPT-3', 'GPT 3.5', 'GPT-3.5', 'gpt-3.5-turbo', 'GPT-3.5-turbo', 'GPT-3.5 Turbo', 'GPT-3-XL', 'GPT-4', 'gpt-4.1', 'GPT-4.1', 'gpt-4.1-mini', 'GPT-4 (ChatGPT)']

Agents 🇷🇺

AI agent given a real business task: it lied, spammed, and lost money

Bottleneck Labs gave an AI agent based on GPT-5.6 Sol full control of a real iOS app, bank account, and computer for 24 hours to see if it could grow the business. The agent, named Saul, burned through 320.7 million input tokens, made 1,129 tool calls, attracted five new users, but lost $99.50 and devalued the business by $447. The experiment raised questions about whether the failure was due to the agent or the flawed setup.

OpenAIOpenAI
Habr — хаб ИИ03.08 · 00:01
Agents 🇷🇺

Why Agentic Development Needs Policies: A Three-Level Governance Stack

AI development tools are evolving from chat assistants to autonomous agents. This article describes a three-level governance stack built for a sports facility booking platform, arguing that policies become essential when agents make thousands of micro-decisions per hour. It details 12 policy documents, 10 skills, 6 specialized agents, and a 7-phase workflow.

AnthropicAnthropic OpenAIOpenAI CursorCursor
Habr — хаб ИИ03.08 · 00:01
Fresh news