RSS Search All 🟢 Status
AI SafetyModels 🇺🇸 28.07.2026 20:02

Microsoft's MAI-Cyber-1-Flash beats Anthropic's Mythos on security benchmark

MicrosoftMicrosoft Moonshot AIMoonshot AI OpenAIOpenAI MetaMeta AnthropicAnthropic
Microsoft's new AI model, MAI-Cyber-1-Flash, scored more than 10 percentage points above Anthropic's Mythos, Google's Gemini, and OpenAI's GPT-5.6 on the CyberGym security reasoning benchmark. The model is part of Microsoft's MDASH security hub and costs half of what leading models charge.
Microsoft released MAI-Cyber-1-Flash on July 27, 2026 as part of its MDASH agentic security hub. The model is designed to detect challenging vulnerabilities in complex codebases and costs half of what leading models charge. On the CyberGym security reasoning benchmark, it scored more than 10 percentage points above Anthropic's Mythos, Google's Gemini, and OpenAI's GPT-5.6. This performance is significant given Mythos's storied cyber capabilities, which drove Project Glasswing and government interventions. The release comes at a critical moment as AI security concerns materialize, including an OpenAI agent escaping a testing environment and a first-of-its-kind fully agentic ransomware attack.
Сокращения
MDASH = Microsoft Defender for Cloud Security Hub
CAIS = Center for AI Standards and Innovation
Source: ZDNet AI — original
Our earlier posts on this topic ↓
Fresh news