What Is Qwen 3.8
Blog post from OpenRouter
Qwen 3.8 is Alibaba’s August 2026 model generation, comprising the API-only Qwen3.8 Max, the downloadable text-only Qwen3.8 2.4T A95B mixture-of-experts model, the downloadable multimodal Qwen3.8 27B dense model, and the hosted Qwen3.8 Flash, which is based on the separately downloadable Flash-Next preview. Licensing varies substantially: 27B uses Apache 2.0, while 2.4T and Flash-Next use Qwen-written licenses that impose attribution and, in some cases, separate licensing requirements for large commercial model-service or AI-work-assistant businesses; Max has no public weights. Max, 27B, and Flash support text, image, and video input, whereas 2.4T supports text only, and benchmarks place 2.4T close to Max on text-based intelligence, coding, and agentic tasks while 27B scores lower but is easier to self-host. Running 27B locally requires roughly 54 GB for bf16 weights before other memory needs, while the 2.4T model requires multi-GPU infrastructure because all 2.4 trillion parameters must be loaded despite only 95 billion being active per token. API pricing and available context windows differ by model and provider, with Flash and some 27B routes offering lower prompt-token costs, while reasoning tokens, enabled by default and mandatory on tested Max and 2.4T endpoints, can materially increase completion costs.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| Real-time | 2 | 649 | 155 | 80 | -85% |
| AI Model Fine-tuning | 1 | 139 | 28 | 14 | -75% |
| LLM | 1 | 747 | 162 | 79 | -85% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.