Qwen3.8 Max Free Online: Alibaba 2.4T Flagship AI

Qwen3.8 Max

Current version: 3.8 Max updated

Alibaba flagship: 2.4T params, 1M context, multimodal

Category
Language
Modality
Text · Reasoning · Tool use
Context
1M
Released
Aug 2026
Update

What's new in Qwen3.8 Max

  • Context extended to 1,000,000 tokens with output up to 131,072
  • 2.4T MoE parameters versus 1T in the previous generation, about 95B active per request
  • Native multimodality: text, image and video input
  • Terminal-Bench 2.1 score of 86.6 against 51.5 for Qwen3.6 on version 2.0
  • SWE-bench Pro at 67.7 versus 49.5 for Qwen3.6
Strengths

What it's the best tool for

  • 1,000,000-token context fits an entire codebase or a full book
  • 2.4T MoE parameters with 95B active: flagship quality without flagship pricing
  • Multimodal input handling text, images and video in a single request
  • Agentic workflows with autonomous runs lasting more than ten days
  • Terminal-Bench 2.1 at 86.6 and SWE-bench Pro at 67.7, matching top-tier coding models
  • Strong multilingual output that does not read like a translation
Limitations

When to reach for something else

  • Open weights are promised but were not published at launch
  • No full model card or reproducible benchmark configuration yet
  • As a Chinese model it answers cautiously on politically sensitive topics
  • Response latency grows noticeably at the far end of the context window
  • No image or video generation — it only accepts them as input
Sample output

How Qwen3.8 Max responds

Prompt
Here is a 40,000-line repository. Find why the checkout_flow integration test fails, propose a fix and write a unit test for it.
Qwen3.8 Max
Traced the stack: the failure is not in checkout_flow but in the payment_gateway_mock fixture. Root cause: the mock returns status 'pending' while the handler has expected 'authorized' since version 2.3. The contract changed in PR #418 and the fixture was never updated. Fix goes in services/payments/gateway.py:112 — treat 'pending' as a retryable status instead of raising. Unit test covering both statuses is below; it will catch the regression if the contract drifts again.
Where teams use it

Four scenarios where it pays for itself

01
Large codebase work
A 1M-token context takes the whole project at once, no chunking
02
Agentic pipelines
Autonomous multi-day runs: edits, tests, log reading and iteration without a human
03
Document and video analysis
Contracts, reports and meeting recordings parsed in a single request
04
Multilingual content
Natural output across languages with terminology held consistent
About model

More about Qwen3.8 Max

Qwen3.8 Max — Alibaba's Flagship AI Model

Qwen3.8 Max is the most capable model in the Qwen family as of August 2026. It packs 2.4 trillion parameters into a sparse Mixture-of-Experts architecture, activating roughly 95 billion per request. That design keeps latency reasonable for a model of this size while landing well below Western flagships on price.

What Qwen3.8 Max can do

The model is natively multimodal — it accepts text, images and video as input. The context window is 1,000,000 tokens, enough for an entire mid-sized codebase or around 750 pages of text. Maximum output runs to 131,072 tokens, so it can produce a substantial module in a single pass.

Alibaba built this release around agentic work. In internal testing, Qwen3.8 Max sustained autonomous software development for more than ten days straight — writing code, running its own tests, reading logs and iterating without human input. It scores 86.6 on Terminal-Bench 2.1 and 67.7 on SWE-bench Pro.

Multilingual strength

Qwen has always been strong across languages, and this release continues that. It handles documents, code comments and long conversations in dozens of languages without the translated-from-English feel that weaker models produce.

Try it on NetRoom

On NetRoom you can run Qwen3.8 Max straight from the browser — no Alibaba Cloud account, no separate API key, no setup. Billing is per token actually used, so you pay for output rather than a subscription tier.

Who it's for

Developers who need long context and agentic workflows; analysts working through large document sets; teams that care about cost per million tokens rather than the logo on the box.

Versions

Version history of Qwen3.8 Max

Version Date What changed
3.8 Max current
  • Context extended to 1,000,000 tokens with output up to 131,072
  • 2.4T MoE parameters versus 1T in the previous generation, about 95B active per request
  • Native multimodality: text, image and video input
  • Terminal-Bench 2.1 score of 86.6 against 51.5 for Qwen3.6 on version 2.0
  • SWE-bench Pro at 67.7 versus 49.5 for Qwen3.6
3.6 Max Preview
  • First Max release in the 3.6 line, shipped with preview status
  • SWE-bench Verified 73.4 and LiveCodeBench v6 80.4
  • Led six coding and agent benchmarks at launch
  • Available through Alibaba Cloud Bailian and Qwen Studio

Try Qwen3.8 Max
right now

Free access to basic models. No card, no obligations.