Qwen3.8-2.4T Model Surfaces on Hugging Face with Agent Capabilities
multimodalart · x · 2026-08-12
A model page for Qwen3.8-2.4T-A95B has appeared on Hugging Face. The leaked configuration reveals a massive 2.4T total parameter count with 95B active parameters, natively supporting robust vision and Agent tool-calling capabilities.
Developer @multimodalart praised the model as 'really good,' particularly highlighting its vision and agent functionalities, and hinted that a smaller version is coming soon.
Related event: Alibaba's Qwen3.8-Max 2.4T MoE Model Tops Hugging Face Trending(11 posts)→
More from Models
- Qwen 1-bit Quantization Shrinks Model to 397GB, a 91% Reduction — danielhanchen · 2026-08-13
- Users Complain About Claude's Unstoppable Chain-of-Thought Output — GuyHachmon · 2026-08-13
- OpenAI's gpt-live-1 Achieves Near-Perfect Conversational Turn Detection — pbbakkum · 2026-08-13
- xAI Launches Grok 4.6 Across Cursor, API with 2x Token Promo — aman_madaan · 2026-08-13
- xAI Exec Hints at Grok 4.5: Focused on Coding Agents — aman_madaan · 2026-08-13
- Reddit Speculates: Is Grok 4.6 a Fine-tune of Kimi K3? — robertpro01 · 2026-08-13