ModelsBusiness & Market 🇩🇪 01.08.2026 00:08

DeepSeek's new Flash model has matched OpenAI GPT-5.6 Luna at roughly 60% lower cost.

DeepSeekDeepSeek OpenAIOpenAI
DeepSeek has released the budget model V4 Flash "0731", which scored 50 points on the Artificial Analysis Intelligence index, ten points more than the previous version, V4 Flash. The model trails the budget model OpenAI GPT-5.6 Luna by just one point, but costs about 60% less per task, thanks to a 98% discount on cache. Improvements are particularly noticeable in agentic tasks: in the GDPval benchmark, the score rose from 1189 to 1559 Elo points, and hallucinations have become less frequent.
DeepSeek has released a new budget model, V4 Flash "0731", which scores 50 points on the Artificial Analysis Intelligence index, ten points more than the previous version released in April 2026. This is just one point less than OpenAI's budget model GPT-5.6 Luna, while the cost per task is about 60 percent lower, even accounting for OpenAI's 80 percent price cut. The key factor is DeepSeek's 98 percent cache discount, significantly higher than the industry average of 90 percent. The model also consumes twelve percent fewer tokens than its predecessor. Improvements span all test categories, especially agentic tasks: in the GDPval benchmark, which measures complex real-world office work, the score rose from 1189 to 1559 Elo points, and the hallucination rate decreased. The architecture remains unchanged: 284 billion parameters, of which 13 billion are active, with a context window of one million tokens. The model's weights are distributed under the MIT license on Hugging Face.
Сокращения
MIT = Massachusetts Institute of Technology — Массачусетский технологический институт
Source: The Decoder (DE) — original
Our earlier posts on this topic ↓
Fresh news