Alibaba/Qwen

Latest AI news, models and releases from Alibaba/Qwen. ['CodeQwen1.5', 'Cotype Light 3', 'Cotype Pro 3', 'Fun-Realtime-TTS', 'Hanguang', 'Hanguang 800', 'HappyOyster 1.0', 'ICN Switch 1.0', 'not specified', 'Panjiu', 'Qianwen', 'Qoder Security', 'QVQ-72B-Preview', 'QVQ-Max', 'Qwen', 'Qwen 2', 'Qwen2.5', 'Qwen2.5-0.5B', 'Qwen2.5-14B', 'Qwen2.5-14B-Instruct-1M', 'Qwen2.5-1.5B', 'Qwen2.5-32B', 'Qwen2.5-3B', 'Qwen-2.5 72B', 'Qwen2.5-72B', 'Qwen2.5-7B', 'Qwen2.5-7B-Instruct-1M', 'Qwen2.5-Coder', 'Qwen2.5-Coder-0.5B', 'Qwen2.5-Coder-0.5B-Instruct', 'Qwen2.5-Coder-14B', 'Qwen2.5-Coder-1.5B', 'Qwen2.5-Coder-1.5B-Instruct', 'Qwen2.5-Coder-32B', 'Qwen2.5-Coder-32B-Instruct', 'Qwen2.5-Coder-3B', 'Qwen2.5-Coder-3B-Instruct', 'Qwen2.5-Coder-7B', 'Qwen2.5-Coder-7B-Instruct', 'Qwen2.5-Coder-Instruct']

Research 🇨🇳

GSPO: Scalable Reinforcement Learning for Language Models

Alibaba Qwen proposes Group Sequence Policy Optimization (GSPO), a new RL algorithm that addresses instability in existing methods like GRPO. GSPO uses sequence-level optimization to improve training efficiency, stability, and compatibility with Mixture-of-Experts models, contributing to the performance of Qwen3 models.

Alibaba/QwenAlibaba/Qwen
Alibaba Qwen27.07 · 17:05
Models 🇨🇳

Alibaba releases Qwen-Image: a 20B MMDiT model for text-accurate image generation

Alibaba's Qwen team has released Qwen-Image, a 20-billion-parameter MMDiT image foundation model that excels in complex text rendering (including bilingual Chinese-English), consistent image editing, and strong performance across benchmarks like GenEval and DPG. The model supports poster creation, slide generation, and precise text rendering on small regions.

Alibaba/QwenAlibaba/Qwen
Alibaba Qwen27.07 · 17:05
Fresh news