Alibaba/Qwen

Latest AI news, models and releases from Alibaba/Qwen. ['CodeQwen1.5', 'Cotype Light 3', 'Cotype Pro 3', 'Fun-Realtime-TTS', 'Hanguang', 'Hanguang 800', 'HappyOyster 1.0', 'ICN Switch 1.0', 'not specified', 'Panjiu', 'Qianwen', 'Qoder Security', 'QVQ-72B-Preview', 'QVQ-Max', 'Qwen', 'Qwen 2', 'Qwen2.5', 'Qwen2.5-0.5B', 'Qwen2.5-14B', 'Qwen2.5-14B-Instruct-1M', 'Qwen2.5-1.5B', 'Qwen2.5-32B', 'Qwen2.5-3B', 'Qwen-2.5 72B', 'Qwen2.5-72B', 'Qwen2.5-7B', 'Qwen2.5-7B-Instruct-1M', 'Qwen2.5-Coder', 'Qwen2.5-Coder-0.5B', 'Qwen2.5-Coder-0.5B-Instruct', 'Qwen2.5-Coder-14B', 'Qwen2.5-Coder-1.5B', 'Qwen2.5-Coder-1.5B-Instruct', 'Qwen2.5-Coder-32B', 'Qwen2.5-Coder-32B-Instruct', 'Qwen2.5-Coder-3B', 'Qwen2.5-Coder-3B-Instruct', 'Qwen2.5-Coder-7B', 'Qwen2.5-Coder-7B-Instruct', 'Qwen2.5-Coder-Instruct']

Applications 🇷🇺

How to reduce LLM costs by 70 times: experience migrating from Claude Sonnet to Qwen 3.5 Flash

JobPath migrated from expensive Claude Sonnet to cheap Qwen 3.5 Flash for job card generation, cutting monthly costs from ~$2000 to ~$30 while quality dropped slightly from 4.48 to 4.06 on a 1-5 scale. The article details selection via evaluation with a blind judge, issues with cheap models (empty responses, type mismatches, field typos), and error-handling lessons including a billing mistake that wasted 16,000 calls in one night.

AnthropicAnthropic Alibaba/QwenAlibaba/Qwen DeepSeekDeepSeek
Habr — хаб NLP24.07 · 06:03
Fresh news