❯ Alibaba Open-Sources Qwen3.8 Weights: 27B Multimodal Model Runs on Consumer GPUs
HIGHLIGHTSAlibaba’s Tongyi team released the full Qwen3.8 weight suite under the Apache 2.0 license. The headline Qwen3.8-27B is a natively multimodal dense model that the team says outperforms the larger Qwen3.7-Plus overall, with particular strength in real-world coding and office workflows. The flagship Qwen3.8-2.4T-A95B (2.4 trillion total parameters, 95 billion activated) is also open now.
LOCAL-FRIENDLYThe 27B version offers 262K native context tokens, expandable to 1M, and handles images, documents, and long video. Its target hardware is high-end consumer cards with 24GB VRAM; the quantized version runs in 17GB of memory. Qwen open-source models already held the No.1 share in local inference, and over a dozen inference platforms integrated the new release on day one. Developer Simon Willison tested it and called the output quality among the best he has seen in local models.
OPEN-SOURCE RACEIt shared the headlines with Zhipu’s GLM-5.3 on the same day, as two Chinese labs staked out “strongest open-source coder” and “most capable local small model” within 24 hours. Enterprises buying APIs will need to re-draw the line between self-hosting and external procurement; developers building local rigs are redoing their price-comparison tables this week. The inference-cost curve is being pushed down by both labs at once.
▪ SIGNALA 27B model beats its own larger predecessor; capability density is improving faster than parameter stacking.