Alibaba releases Qwen3.8-27B, compact vision-language model with extended context
Alibaba has released Qwen3.8-27B, a 27-billion parameter vision-language model that builds on the Qwen3.5 architecture with improvements in coding, professional work, and agentic task completion. The model features a 262,144 token native context length extensible to 1 million tokens, native support for image and video understanding, and flexible reasoning controls including a thinking mode that can be tuned per request. Qwen3.8-27B introduces architectural innovations including Gated DeltaNet for linear attention alongside traditional gated attention mechanisms, and is designed as a compact, deployable option compared to larger models in the family. The model is available through Hugging Face Transformers and compatible with multiple inference frameworks including vLLM and SGLang, with a managed version coming soon through Qwen Cloud.