qwen3.8-2.4t-a95b · Qwen
Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a 2.4‑trillion‑parameter sparse Mixture-of-Experts (MoE) model with approximately 95 billion active parameters. It is built for autonomous, long‑duration tasks: multi‑day code runs, reproducing research papers, and self‑improvement.
Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a 2.4‑trillion‑parameter sparse Mixture-of-Experts (MoE) model with approximately 95 billion active parameters. It is built for autonomous, long‑duration tasks: multi‑day code runs, reproducing research papers, and self‑improvement.
Qwen3.8 2.4t A95B has a 1,000,000 token context window. It supports up to 131,072 output tokens.
On AIHubMix, Qwen3.8 2.4t A95B costs $2 per million input tokens and $6 per million output tokens. Cached input reads are billed at $0.5 per million tokens.
Qwen3.8 2.4t A95B accepts text input.
Qwen3.8 2.4t A95B supports tool calling, function calling, structured outputs, web search, long context and thinking. Per-protocol parameter support is listed in the capability table on this page.
Qwen3.8 2.4t A95B is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3.8-2.4t-a95b — no other code changes needed.
Qwen3.8 2.4t A95B is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding…
Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…
Wan3.0 Video Prime is Alibaba Cloud’s preview high-speed edition of its All-in-One video…
Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by…
Qwen Image 3.0 Pro (qwen-image-3.0-pro) is Alibaba Cloud Qwen’s flagship image generation…
Qwen 3.8 Max(qwen3.8-max) is Alibaba Cloud’s flagship native vision-language model, built…
Use Qwen3.8 2.4t A95B via the AIHubMix unified API — one interface for every major LLM.