Home / Companies / OpenRouter / Blog / Post Details
Content Deep Dive

What Is Qwen 3.8

Blog post from OpenRouter

Post Details
Company
Date Published
Author
OpenRouter
Word Count
3,198
Company Posts That Month
26
Language
English
Hacker News Points
-
Post removed?
No
Summary

Qwen 3.8 is Alibaba’s August 2026 model generation, comprising the API-only Qwen3.8 Max, the downloadable text-only Qwen3.8 2.4T A95B mixture-of-experts model, the downloadable multimodal Qwen3.8 27B dense model, and the hosted Qwen3.8 Flash, which is based on the separately downloadable Flash-Next preview. Licensing varies substantially: 27B uses Apache 2.0, while 2.4T and Flash-Next use Qwen-written licenses that impose attribution and, in some cases, separate licensing requirements for large commercial model-service or AI-work-assistant businesses; Max has no public weights. Max, 27B, and Flash support text, image, and video input, whereas 2.4T supports text only, and benchmarks place 2.4T close to Max on text-based intelligence, coding, and agentic tasks while 27B scores lower but is easier to self-host. Running 27B locally requires roughly 54 GB for bf16 weights before other memory needs, while the 2.4T model requires multi-GPU infrastructure because all 2.4 trillion parameters must be loaded despite only 95 billion being active per token. API pricing and available context windows differ by model and provider, with Flash and some 27B routes offering lower prompt-token costs, while reasoning tokens, enabled by default and mandatory on tested Max and 2.4T endpoints, can materially increase completion costs.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Real-time 2 649 155 80 -85%
AI Model Fine-tuning 1 139 28 14 -75%
LLM 1 747 162 79 -85%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.