Alibaba Open-Sources Qwen2.5-VL-32B: Smarter and Lighter
Alibaba has open-sourced Qwen2.5-VL-32B-Instruct, a vision-language model with 32 billion parameters, under Apache 2.0. It outperforms Mistral-Small-3.1-24B and Gemma-3-27B-IT, and even surpasses the larger Qwen2-VL-72B-Instruct on multimodal reasoning benchmarks like MMMU and MathVista. The model is optimized via reinforcement learning for better human alignment and mathematical reasoning.
Alibaba/Qwen
