FEATUREDAI HOT (Curated Pool)· aihot-apiZH15:02 · 08·14
→Qwen releases Qwen3.8 series: a 27B dense multimodal model and open weights for a 2.4T-A95B Max variant
Qwen delivered on its open-source promise with the Qwen3.8 series. Qwen3.8-27B is a natively multimodal dense model that beats Qwen3.7-Plus at only 27B parameters, supports 262K context natively and up to 1M tokens via YaRN, under Apache 2.0. Open weights for the Max-tier Qwen3.8-2.4T-A95B are also available. The post doesn't cover training data, inference cost, or release timeline details.
#Multimodal#Alibaba Qwen#Open source
why featured
Featured · importance 88 · hook + knowledge + resonance
editor take
Qwen3.8-27B beats Qwen3.7-Plus on benchmarks at 27B params, with native multimodal, 262K context, and Apache 2.0 license.
sharp
The parameter efficiency here is wild: a 27B dense model that claims to beat Qwen3.7-Plus across benchmarks. Native vision and document understanding, 262K context window that stretches to 1M tokens with YaRN — that's the entire Three-Body Problem trilogy in one go. Apache 2.0 means no commercial headaches. They also dropped open weights for the Max-tier Qwen3.8-2.4T-A95B. This is just the official tweet though — no training data details, no inference cost, no concrete release timeline. I'd hold off on those numbers. But a 27B model with native multimodal and super-long context is genuinely easier to deploy than anything in the 70B+ range.
HKR breakdown
hook ✓knowledge ✓resonance ✓