Skip to content
All master reviews

Review Machine · Master Review · AI models

Qwen3.8-2.4T-A95B

The charts are the launch, the download is the checkpoint

The open checkpoint behind Qwen3.8-Max is documented as text-only with thinking always on, frozen at its August state while the hosted model took a September post-training pass, and no independent evaluation of the weights themselves was readable on 20 September 2026.

Open weights published 12 August 2026, Hugging FaceArchitecture 2.4T total / 95B active, 92 layers, 512 expertsHosted price $2 in / $6 out per 1M tokens, QwenCloudLicence Qwen3.8-Max LicenseRead 20 September 2026Skip to the verdictRead as text

Verdict age

ProvisionalVendor pages settled · independent runs thin

The weights are five weeks old as we read, and the hosted model took a September post-training pass that has no open checkpoint, so a later checkpoint or a first independent run would change this reading.

  • Early · editors only
  • Provisional · editors in, owners thin
  • Settled · owner reviews read over months

From 19 sources to one verdict

Seven levels, in the same order in every Master Review. Each line is the takeaway; open a level for the evidence under it, or stop when you have enough.

What we read

19 sources, 3 Aug – 20 Sep 2026. Forum and buyer tiers thin. Editorial tier empty.

By tier, heaviest first

T1Official5

The model card, the licence file, the launch post, the QwenCloud page and the repository, read 20 September 2026.

T2Editorial1None

One page, and it scores a September build of the hosted model rather than these weights.

T2News0None readable
T3Buyers1Thin

Counts from one unauthenticated GitHub REST API request, not ratings.

Forum4Thin

Four threads from one Hacker News API search that answered 94 stories.

Other8

Eight write-ups about the release, read 20 September 2026; several sit on vendor-adjacent blogs, so their claims are reported as theirs.

Date window

3 Aug 2026 to 20 Sep 2026

Editorial reviews 20 Sep 2026Official pages read 20 September 2026; the launch post is dated 3 August 2026.

Couldn’t read, so not used

  • South China Morning Post report on the releasepaywall
  • Qwen3.8-Max Overview page on the forum announcementno dated page of its own could be opened

Nothing that rests only on a source we could not read is counted, and the paywalled report is not cited.

Read, then rejected

  • Thomas Wiegold's hands-on review · a different version, the July preview, and it reports running the model
  • Trilogy AI's StackPerf run · a different version, one matched run
  • DataCamp's benchmark breakdown · restates the maker's table for the hosted build

What they measured

2.4 trillion parameters with 95 billion active per token, a native 262,144-token context, and 4,142 stars on the repository five weeks after the weights landed.

Parameters

2.4trilliontotal, with 95 billion active per token, 92 layers and 512 experts

Hugging Face · read 20 Sep

Native context

262,144tokensthe card says it is extensible to 1,010,000

Hugging Face · read 20 Sep

Weight file size

4.9TB BF16the FP8 build is put near 2.4 TBA provider's own write-up, not the card

OrcaRouter · read 20 Sep

Repository figures

4,142starswith 314 forks and 20 open issues

GitHub REST API · read 20 Sep

Terminal-Bench 2.1

Vendor table. The card says its own entries ran on a coding harness while the rivals' cells are their best published scores.

Output speed

38.6tokens/sthe independent page calls this slow; it is the September hosted build, not the weightsDifferent version

Artificial Analysis · read 20 Sep

Where they agree

The checkpoint is the Max-class core, documented consistently, with no independent run of the weights themselves.

One tick per editorial review, in order of publication:AA Artificial Analysis

The downloadable checkpoint is the Max-class core

0/1

The card, the launch post and seven write-ups give the same architecture and the same limits. T1Official Other

The licence grant is broad below two revenue lines

0/1

Use, modification, hosting and fine-tuning are granted free of charge; the conditions begin at 100 million monthly users or US$20 million monthly revenue, and at US$50 million for Model-as-a-Service and AI-work-assistant businesses. T1Official Other

The front group, on the maker's own table

0/1

Terminal-Bench 2.1 at 86.6 and PaperBench at 93.0, both vendor figures whose footnotes say Qwen's own entries ran on a coding harness while the rivals' cells are published scores from elsewhere, and one audit that reads the model as belonging in the frontier cohort. T1Official Other

A fixed artifact you can pin

0/1

Two write-ups put custody and reproducibility first as the reason to take the weights over a rolling endpoint. Other

Attention, which says little on its own

0/1

4,142 stars, 314 forks and 20 open issues on the repository on 20 September 2026, and one API search that answered 94 stories about the family. T3Buyers Forum

What keeps coming up against it

The version you can download is not the version being measured, and the licence is the part with the longest tail.

The version you can download is not the version being measured

Well documented · count not published
Documented byT1Hugging Face model cardT1QwenCloudWrite-upsHow commonUnknown: no source we could read counts it; a provider's write-up states that no independent benchmark of the downloadable checkpoint has been published
T1Hugging Face model card

The card describes the checkpoint as text-only with thinking required for all interactions, and points to the hosted version for vision, a non-thinking mode, a 1M window and built-in tools.

  • Hugging Face a text-only model that requires thinking mode for all interactions

Hugging Face · read 20 Sep

T1QwenCloud

The hosted page lists image and video input, function calling, structured outputs and a 1M context, none of which the checkpoint's card claims.

QwenCloud · read 20 Sep

Write-ups

Three write-ups state the same split, one of them describing the hosted product as the official version built on these weights with more features.

OrcaRouter · read 20 SepAI Tools Recap · read 20 SepGroundy · read 20 Sep

The weights carry a separate licence with revenue conditions

0/1

The shipped file is titled the Qwen3.8-Max License, not the Apache 2.0 of the previous generations, and six write-ups report the same. T1Official Other

Serving the checkpoint is a rack-scale job

0/1

BF16 weights put near 4.9 TB, an FP8 build near 2.4 TB, and a serving floor the write-ups place at multi-node accelerators. Other

Thinking cannot be switched off in the download

0/1

The card states that thinking cannot be disabled, and the hosted version offers a non-thinking mode. T1Official Other

Where they split

Five places the evidence pulls two ways, starting with the date the weights landed.

The date the weights landed

Hacker News

The thread announcing the checkpoint is dated 12 August 2026.

One write-up

Against that: Dates the open release 13 August 2026.

Our readWe take 12 August, the date on the checkpoint's own thread.

The licence the repository shows against the licence the weights carry

T1GitHub

The repository is listed under Apache 2.0 on 20 September 2026.

T1The shipped file

Against that: The licence shipped with the weights is the Qwen3.8-Max License, with two revenue conditions.

Our readBoth accurate about different things: the repository's file covers the series, the shipped licence governs the weights.

Whether the open checkpoint matches the hosted model

Groundy

Reports the independent index scoring the hosted build and the open checkpoint at the same aggregate figure.

A provider's write-up

Against that: States that no independent benchmark of the downloadable checkpoint has been published, and that the hosted build carries a September post-training the weights do not.

Our readA matching aggregate on one index does not make two artifacts one product, and the version gap is most of the difference.

What the launch table measures against

T1Model card footnotes

Its own entries ran on a coding harness, while the rivals' cells are their best published scores from elsewhere.

T2Artificial Analysis

Against that: Scores the September hosted build at 45 on its intelligence index, its own number for a different version.

Our readNeither is a head-to-head, and no evaluation we could read scores the checkpoint itself.

Where the developer threads went

Hacker News, August

The launch post drew 1,124 points and 612 comments, the checkpoint 713 points and 171.

Hacker News, September

Against that: Among the stories the one search returned, the release thread after the checkpoint is the Flash-Next one, dated 26 August 2026, and the September threads are about the smaller sibling's quantisations.

Our readThe flagship's day was the launch, and the weeks after belong to the models people can run.

Who it's for, who should pass

For a team that needs the weights in its own custody and reads the release for what it is. Pass if you wanted the model the hosted version is.

It suits you if

  • You need the weights in your own custody, on hardware you already operate.The licence covers hosting and fine-tuning below its revenue lines, and a fixed checkpoint can be audited and pinned.
  • Your work is long-horizon agentic coding and text input is enough.The card and the launch post are written around coding and long tasks, and the one performance figure we could read puts the hosted build at 38.6 output tokens per second and calls that slow.
  • You are reading the release rather than shopping for a product.A documented checkpoint with a published licence is worth studying on its own terms, even where the product around it is not.

Pass if

  • You wanted the model the hosted version is.The card describes the weights as text-only with thinking always on, while the hosted product adds image and video input, a non-thinking mode, a 1M window and built-in tools.
  • You have no multi-node GPU capacity.Three write-ups put the serving floor at rack scale and recommend the smaller sibling for one machine.
  • You need unconditional open weights.The previous generations shipped under Apache 2.0; these weights carry the Qwen3.8-Max License, and the file itself is the operative text.

This is a reading of published reviews, not medical, financial or legal advice. For a decision about your health, your money or your rights, a qualified professional is the right next step, and not a review.

The verdict

Vendor documentation settled, no independent run of the weights, and a version gap that is the whole story.
VerdictProvisionalVendor pages settled · independent runs thin

A genuine Max-class checkpoint with a documented architecture and a permissive grant, published under a licence with revenue conditions, and frozen at its August state while the hosted model moved on, with no independent evaluation of the weights themselves readable.

Confidence, by tier

Officialstrong

The card, the licence file, the launch post, the model page and the repository, all read on 20 September 2026.

Editorialnone

The one evaluation page we could open scores a September build of the hosted model, not these weights.

Forumthin

Four threads from one API search that answered 94 stories, weighted to the launch week.

Buyersthin

One GitHub API call for stars, forks and open issues, which measure attention and not quality.

Editorial evidence
under a month old
Newest report
None read
Read
20 Sep · month 1

Share line

You can download the max-class weights. The max-class model stays on the API.
Review MachineQwen3.8-2.4T-A95B · Master Review

Rests onThe model card describes the checkpoint as text-only with thinking that cannot be disabled, against the hosted version's vision, non-thinking mode and 1M window, and a provider's write-up of 20 September 2026 reports that the September post-training has no open checkpoint.

Sources

19 sources, heaviest tier first. Every figure above comes from one of these.

  1. T1OfficialQwen3.8-Max: A New Bar for Coding and Cowork3 Aug
  2. T1OfficialQwen3.8-2.4T-A95B model cardread 20 Sep
  3. T1OfficialQwen3.8-Max License, shipped with the weightsread 20 Sep
  4. T1OfficialQwen3.8-Max model pageread 20 Sep
  5. T1OfficialQwenLM/Qwen3.8 repository and READMEread 20 Sep
  6. T2EditorialQwen3.8 Max (0902)read 20 Sep
  7. T3BuyersRepository figures for QwenLM/Qwen3.8one unauthenticated request for stars, forks and open issuesread 20 Sep
  8. ForumQwen3.8-Max: A New Bar for Coding and Cowork, 1,124 points, 612 comments3 Aug
  9. ForumQwen3.8-2.4T, 713 points, 171 comments12 Aug
  10. ForumQwen3.8-Flash-Next, 704 points, 233 commentsa different release of the same family, cited only in Where reviewers split26 Aug
  11. ForumSearch for Qwen3.8, 94 storiesone search request that reports a total, no pagingread 20 Sep
  12. OtherQwen3.8 Max release auditread 20 Sep
  13. OtherQwen3.8-Max against the open weightspublished by a provider that resells model accessread 20 Sep
  14. OtherOpen weights under a custom licenceread 20 Sep
  15. OtherTwo licences for one releaseread 20 Sep
  16. OtherThe first Max-class open weightsread 20 Sep
  17. OtherThe custom licence decodedread 20 Sep
  18. OtherWeights against the APIread 20 Sep
  19. OtherQwen 3.8 review and hardwareread 20 Sep

MethodOn 20 September 2026 we read 19 sources dated 3 August to 18 September 2026, from the model card, the licence file and the launch post to one independent evaluation page, eight write-ups and four Hacker News threads, and synthesised them with AI. Nothing was downloaded, run or tested.

Dates without a year are 2026.

The review as text

The same review as one piece of writing · 7 min read

Qwen3.8-2.4T-A95B is the open-weight checkpoint behind Qwen3.8-Max, Alibaba's 2.4-trillion-parameter mixture-of-experts flagship, and the first Qwen-Max-class model the company has published as downloadable weights. The launch post went up on 3 August 2026 and promised the weights "next week". The checkpoint reached Hugging Face on 12 August 2026, while the hosted Max-class product stayed on sale through QwenCloud with features the download does not carry.

This reading covers nineteen sources read on 20 September 2026: five pages published by Alibaba or Hugging Face, one independent evaluation page, eight write-ups, four Hacker News threads from one API search that answered 94 stories, and one GitHub API call for the repository's figures. The material dates from 3 August to 18 September 2026. No independent evaluation of these weights was readable, and the news report we found was paywalled, so both gaps are named below rather than papered over.

Consensus

What the sources agree on is what the checkpoint is, not how well it works. Its shape and limits are documented consistently: 2.4 trillion parameters in total with 95 billion active per token, 92 layers, 512 experts, a native context of 262,144 tokens that the card says can be extended to 1,010,000, and text as its only input. Nine of the nineteen sources, two official pages and seven write-ups, restate that description. On capability the agreement stops. The only independent evaluation we could open scores a September build of the hosted model, not these weights, and no source we read claims to have run the checkpoint itself. Confidence here is thin, and the reason is the version boundary rather than a disagreement between sources.

Recurring strengths

The grant is broad below its two revenue lines. The licence file shipped with the weights grants use, modification, publication, distribution, sale, hosting, fine-tuning and derivative works free of charge. Seven of the nineteen sources, that licence file and six write-ups, describe it that way, and the conditions bite only at scale. Attribution is required above 100 million monthly active users or US$20 million monthly revenue, and a separate licence above US$50 million for Model-as-a-Service and AI-work-assistant businesses.

The download is an artifact you can pin. Two of the write-ups put custody and reproducibility first, the practical case for a versioned checkpoint over a rolling endpoint.

The maker's table puts the family in the front group. The card's Terminal-Bench 2.1 row reads 86.6 for Qwen3.8-Max against 88.8 for GPT-5.6 Sol and 84.6 for Claude Opus 4.8, and its PaperBench row, 93.0, is the highest in that table. Three of the nineteen sources, the card, the launch post and one audit, place the Max-class build there. Its footnotes say its own entries ran on a coding harness while the rivals' cells are published scores from elsewhere.

Attention is easy to measure and says little. The repository held 4,142 stars with 314 forks and 20 open issues on 20 September 2026, five weeks after the weights landed, and one API search answered 94 stories about the family, weighted to the launch week.

Recurring complaints

A separate licence, not the Apache 2.0 of the previous generations. Seven of the nineteen sources, the licence file and six write-ups, report that the weights carry a document titled the Qwen3.8-Max License, and the write-ups read its second clause as aimed at providers who resell access. Whether that licence permits a reader's own use is a legal question this piece does not answer.

The serving floor is a rack. Three of the write-ups put the BF16 weights near 4.9 TB, an FP8 build near 2.4 TB, and the smallest documented serving configuration at multi-node accelerators, which is why the same write-ups recommend the smaller sibling instead.

Thinking cannot be switched off in the download. Four of the nineteen sources, the model card and three write-ups, report that every response begins with reasoning and that the mode cannot be disabled, while the hosted version offers a non-thinking mode.

This is a reading of published reviews, not medical, financial or legal advice. For a decision about your health, your money or your rights, a qualified professional is the right next step, and not a review.

Where reviewers split

The date the weights landed. The thread announcing the checkpoint is dated 12 August 2026, and one write-up dates the release 13 August. We take 12 August, the checkpoint's own thread date.

The licence the repository shows against the licence the weights carry. The GitHub repository is listed under Apache 2.0, and the file shipped with the weights is the Qwen3.8-Max License. Both are official pages read the same day and accurate about different things. Our read: the repository's file covers the series, the shipped licence governs the weights.

Whether the open checkpoint matches the hosted model. One audit reports that the independent index scored the hosted build and the open checkpoint at the same aggregate figure. A provider's write-up states that no independent benchmark of the downloadable checkpoint has been published, and that the hosted build carries a September post-training the weights do not. Both true: a matching aggregate on one index does not make two artifacts one product, and the version gap is most of the difference.

What the launch table measures against. The card's footnotes say its entries ran on one harness and the rivals' cells elsewhere. The independent page we could open scores the September hosted build at 45 on its intelligence index, its own number for a different version. Neither is a head-to-head, and no evaluation we could read scores the checkpoint itself.

Where the developer threads went. The launch-post thread drew 1,124 points and 612 comments, and the checkpoint's own thread 713 points and 171. Among the twenty stories that search returned, the later ones are about the smaller sibling's quantisations and the Flash-Next release. The flagship's day was the launch, and the weeks after belong to the models people can actually run.

Who it suits

You need the weights in your own custody, on hardware you already operate. The licence covers hosting and fine-tuning below its revenue lines, and a fixed checkpoint is the thing that can be audited and pinned.

Your work is long-horizon agentic coding and text input is enough. The card and the launch post are both written around coding and long tasks, and the one performance figure we could read is a warning about patience: the independent page puts the hosted build at 38.6 output tokens per second and calls that slow.

You are reading the release rather than shopping for a product. A documented checkpoint with a published licence is worth studying on its own terms.

Who should pass

You wanted the model the hosted version is. The card describes the weights as text-only with thinking always on. The hosted product adds image and video input, a non-thinking mode, a 1M window and built-in tools, and five of the eight write-ups say the two are not the same artifact.

You have no multi-node GPU capacity. Three of the write-ups put the serving floor at rack scale and recommend the smaller sibling for one machine.

You need unconditional open weights. Previous Qwen generations shipped under Apache 2.0. These weights do not, and the file itself is the operative text. One verdict a week: the most useful Master Review we finished, the complaint that kept appearing, and who should skip it. Get the weekly verdict.

Sources

  1. Qwen3.8-2.4T-A95B model card, official, read 20 September 2026.
  2. Qwen3.8-Max License, shipped with the weights, official, read 20 September 2026.
  3. Qwen3.8-Max: A New Bar for Coding and Cowork, Qwen blog, official, published 3 August 2026.
  4. Qwen3.8-Max, QwenCloud model page, official, read 20 September 2026.
  5. QwenLM/Qwen3.8 repository, official, read 20 September 2026.
  6. Repository figures for QwenLM/Qwen3.8, the public GitHub REST API, one request, 20 September 2026.
  7. Qwen3.8 Max (0902), Artificial Analysis, independent evaluation, read 20 September 2026.
  8. Hacker News API search for Qwen3.8, 94 stories, read 20 September 2026.
  9. Hacker News, the launch post, forum, 3 August 2026.
  10. Hacker News, the checkpoint, forum, 12 August 2026.
  11. Hacker News, the Flash-Next release, forum, 26 August 2026.
  12. Groundy, Qwen3.8 Max release audit, other, read 20 September 2026.
  13. OrcaRouter, Qwen3.8-Max against the open weights, other, read 20 September 2026.
  14. Machine Brief, open weights under a custom licence, other, read 20 September 2026.
  15. SQ Magazine, two licences for one release, other, read 20 September 2026.
  16. The Ledger, the first Max-class open weights, other, read 20 September 2026.
  17. Latent East, the custom licence decoded, other, read 20 September 2026.
  18. AI Tools Recap, weights against the API, other, read 20 September 2026.
  19. AirMore, Qwen 3.8 review and hardware, other, read 20 September 2026.

On 20 September 2026 we read 19 sources dated 3 August to 18 September 2026, from the model card, the licence file and the launch post to one independent evaluation page, eight write-ups and four Hacker News threads, and synthesised them with AI. Nothing was downloaded, run or tested.

One verdict a week.

Every week, the most useful Master Review we finished: what the internet agrees on, the complaint that kept appearing, and who should skip it.

By subscribing you agree to receive one weekly email from Review Machine. You can unsubscribe at any time.