Hugging Face

Latest AI news, models and releases from Hugging Face. ['Diffusers', 'GLM 5.2', 'LeRobot v0.6.0', 'Nunchaku Lite', 'qwen3-32b', 'Qwen/Qwen3-4B', 'Qwen/Qwen3.5-122B-A10B', 'Qwen/Qwen3.5-4B', 'SmolLM', 'SmolLM2-1.7B', 'SmolLM3', 'SmolLM3 3B', 'SmolLM3-3B', 'SmolVLM-256M', 'TinyLlama', 'tokenizers', 'Transformers v4', 'Transformers v5', 'TRL', 'Xenova/distilbert-base-uncased-finetuned-sst-2-english', 'Xenova/whisper-tiny.en']

Open Source 🇺🇸

Hugging Face and Microsoft Foundry: Open Models on Managed GPUs

Hugging Face and Microsoft Foundry have partnered to bring open-weight models from the Hugging Face ecosystem to the Foundry platform, providing managed GPU compute. The curated Hugging Face Collection offers security-screened models across modalities, with deployment templates and automated runtime management. This enables enterprises to deploy open models with the same ease as proprietary models on Foundry.

MicrosoftMicrosoft Hugging FaceHugging Face OpenAIOpenAI AnthropicAnthropic MetaMeta MistralMistral DeepSeekDeepSeek
Hugging Face blog27.07 · 04:06
Open Source 🇺🇸

Hugging Face Transformers Backend for vLLM Achieves Native Performance

The integration of Hugging Face Transformers as a modeling backend in vLLM now matches or exceeds native vLLM throughput across various Qwen3 models, including dense and mixture-of-experts architectures, without requiring custom porting. This is achieved through dynamic layer fusions at runtime using torch.fx and AST manipulation, allowing any compatible model to run at native speed with a single flag.

Hugging FaceHugging Face
Hugging Face blog27.07 · 04:05
Fresh news