<!-- mobian-agent-page publisher="dailydev" canonical="https://daily.dev/posts/qwen3-8-max-alibaba-s-2-4-trillion-parameter-open-weight-model-benchmarked-0pqptx1df" -->

---
title: Qwen3.8 Max: Alibaba&#x27;s 2.4 trillion parameter...
description: Alibaba&#x27;s Qwen team has announced Qwen3.8 Max, a 2.4 trillion parameter mixture-of-experts model positioned as frontier-competitive. A preview is live but open...
canonical: https://daily.dev/posts/qwen3-8-max-alibaba-s-2-4-trillion-parameter-open-weight-model-benchmarked-0pqptx1df
twitter:card: summary_large_image
twitter:site: @dailydotdev
og:type: website
og:site_name: daily.dev
og:title: Qwen3.8 Max: Alibaba&#x27;s 2.4 trillion parameter open-weight model benchmarked | daily.dev
og:description: Alibaba&#x27;s Qwen team has announced Qwen3.8 Max, a 2.4 trillion parameter mixture-of-experts model positioned as frontier-competitive. A preview is live but open...
og:url: https://daily.dev/posts/qwen3-8-max-alibaba-s-2-4-trillion-parameter-open-weight-model-benchmarked-0pqptx1df
og:image: https://api.daily.dev/og/posts/0pqptX1DF.png
og:image:alt: Qwen3.8 Max: Alibaba&#x27;s 2.4 trillion parameter open-weight model benchmarked
og:image:width: 1200
og:image:height: 630
og:locale: en
---

> ## Documentation Index
> Fetch the complete documentation index at: https://daily.dev/llms.txt
> Use this file to discover all available pages before exploring further.

# Qwen3.8 Max: Alibaba's 2.4 trillion parameter open-weight model benchmarked

**[Collections](https://daily.dev/sources/collections)** · 3 min read · 56 upvotes · 9 comments

## Summary

Alibaba's Qwen team has announced Qwen3.8 Max, a 2.4 trillion parameter mixture-of-experts model positioned as frontier-competitive. A preview is live but open weights aren't released yet. Across 8 benchmark tasks including 3D rendering, SVG generation, math reasoning, and agentic workflows, it scored 65/80 (81.25%), placing second behind Fable 5 and ahead of Claude Opus 4. It achieved perfect scores on three tasks. A notable limitation is inconsistent behavior during long-running agentic sessions, particularly with Claude Code. The open weights release is confirmed but pending.

## Content

Alibaba has previewed Qwen 3.8 Max, a 2.4 trillion parameter model it's calling second only to Anthropic's Fable 5. The model is available now through Alibaba's Token Plan, Qoder, QoderWork, and Qwen Chat. Open weights are confirmed but no release date or license terms have been given.

## The claim vs. the evidence

The announcement landed a few days after rival Moonshot released Kimi K3 — complete with a benchmark table, pricing, and a hard open-weight date of July 27. Alibaba's launch, by contrast, was a single post on X with no benchmark scores, no model card, and no architecture details.

AI analyst Julien Simon has pointed out the obvious problem: an undated, unverifiable ranking claim functions more as a narrative play than a model launch. There's also a structural awkwardness here — Alibaba holds roughly a 36% stake in Moonshot, so publishing benchmarks that directly compare the two would be uncomfortable either way.

That said, Qwen's two previous flagships shipped API-only, quietly breaking the lab's open-weight tradition. Promising open weights for 3.8 Max is at least a signal they're aware of the optics.

## Independent testing

One detailed third-party evaluation tested Qwen 3.8 Max across eight tasks: 3D rendering, SVG generation, math reasoning, and agentic workflows. It scored 65/80 (81.25%), placing second overall — behind Fable 5 and ahead of Claude Opus 4. It got perfect scores on a bow-and-arrow game, a hard permutation problem, and a long-horizon fine-tuning task.

Separately, AI/ML API ran a Three.js benchmark where each model was given a single file containing a war galley, a Trojan horse, and a Greek war helmet, then asked to build three scenes in one shot. Qwen 3.8 Max finished considerably faster than the other models with no visible drop in output quality.

The one consistent weakness flagged in early testing was unreliable behavior during long agentic sessions in Claude Code. A subsequent update appears to have addressed this, along with improvements to front-end code generation and tool calling reliability.

## Pricing and access

The model is currently available at 10% of standard pricing on Alibaba's Token Plan. On Qoder, a 90% discount drops the billing coefficient to 0.05x, and off-peak hours (10pm–8am Singapore time) reduce it further to 0.01x — roughly 50x more usage for the same spend. A 14-day pro trial with 300 credits is also available.

Qoder also quietly launched a mystery model called "Cantus" priced at 3.2x the standard billing coefficient — about six times more expensive than Qwen 3.8 Max at standard rates — with no benchmarks or lab attribution. Speculation ranges from an unreleased Qwen Ultra variant to a frontier model from another lab being tested under a codename.

Qwen 3.8 Max is now free on Qwen Studio, though the free tier for Qwen Code CLI hosted inference has been discontinued.

## What's still missing

The open weights. Until those ship, the "second only to Fable 5" claim stays unverifiable. The model's predecessor, Qwen3.7-Max, was the highest-ranked Chinese model on ECI, so a 2.4T successor being competitive at the frontier isn't implausible — but implausible and unverified are different problems.

## Community discussion

Top comments from developers on daily.dev.

**@hidden\_rahatkhan** · 6 upvotes

> Even selling my two kidneys won't be enough to run it locally, I guess!

**@geekluffy** · 2 upvotes

> Beating Claude Opus 4 on benchmarks as an open weight model is wild but the note on inconsistent performance during long claude code sessions shows once again that multi turn agentic stability is the real benchmark developers actually care about

**@amizzo** · 0 upvotes

> Closed models will always be superior - just logically. There's no reason to open source something that's superior. The second a developed model becomes superior, they'll close it. Such is capitalism.

**@simos** · 0 upvotes

> What is seems to me with all those criticism over chinese models is that Americans keep to say it is not safe/legit/propaganda/propaganda/ etc etc to save a really big business for them. This is not safe, not an open weight model, baby!!!

---

Tags: [#ai](https://daily.dev/tags/ai), [#llm](https://daily.dev/tags/llm), [#alibaba](https://daily.dev/tags/alibaba), [#mixture-of-experts](https://daily.dev/tags/mixture-of-experts), [#qwen](https://daily.dev/tags/qwen)

[View this post on daily.dev](https://daily.dev/posts/qwen3-8-max-alibaba-s-2-4-trillion-parameter-open-weight-model-benchmarked-0pqptx1df)

```json
{"@context":"https://schema.org","@graph":[{"@type":"Organization","@id":"https://daily.dev/#organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180},"sameAs":["https://twitter.com/dailydotdev","https://github.com/dailydotdev","https://www.linkedin.com/company/daily-dev-ltd"]},{"@type":"WebSite","@id":"https://daily.dev/#website","url":"https://daily.dev","name":"daily.dev","publisher":{"@id":"https://daily.dev/#organization"},"potentialAction":{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https://daily.dev/search?q={search_term_string}"},"query-input":"required name=search_term_string"}}]}
{"@context":"https://schema.org","@type":"TechArticle","headline":"Qwen3.8 Max: Alibaba's 2.4 trillion parameter open-weight model benchmarked","url":"https://daily.dev/posts/qwen3-8-max-alibaba-s-2-4-trillion-parameter-open-weight-model-benchmarked-0pqptx1df","mainEntityOfPage":{"@type":"WebPage","@id":"https://daily.dev/posts/qwen3-8-max-alibaba-s-2-4-trillion-parameter-open-weight-model-benchmarked-0pqptx1df"},"datePublished":"2026-07-19T12:37:43.318Z","dateModified":"2026-07-23T12:21:30.630Z","description":"Alibaba's Qwen team has announced Qwen3.8 Max, a 2.4 trillion parameter mixture-of-experts model positioned as frontier-competitive. A preview is live but open...","image":"https://i.ytimg.com/vi/_fKg3apyWhc/sddefault.jpg","thumbnailUrl":"https://i.ytimg.com/vi/_fKg3apyWhc/sddefault.jpg","isAccessibleForFree":true,"articleSection":"Collections","inLanguage":"en","publisher":{"@type":"Organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180}},"author":{"@type":"Organization","name":"Collections","logo":"https://media.daily.dev/image/upload/s--fk_6ycEi--/f_auto,q_auto/v1780996001/logos/collections?_a=BAMAMiWQ0","url":"https://daily.dev/sources/collections"},"commentCount":9,"discussionUrl":"https://daily.dev/posts/qwen3-8-max-alibaba-s-2-4-trillion-parameter-open-weight-model-benchmarked-0pqptx1df","interactionStatistic":[{"@type":"InteractionCounter","interactionType":{"@type":"LikeAction"},"userInteractionCount":56},{"@type":"InteractionCounter","interactionType":{"@type":"CommentAction"},"userInteractionCount":9}],"keywords":"ai,llm,alibaba,mixture-of-experts,qwen","timeRequired":"PT3M"}
{"@context":"https://schema.org","@type":"BreadcrumbList","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https://daily.dev"},{"@type":"ListItem","position":2,"name":"Collections","item":"https://daily.dev/sources/collections"},{"@type":"ListItem","position":3,"name":"Qwen3.8 Max: Alibaba's 2.4 trillion parameter open-weight model benchmarked"}]}
{"@context":"https://schema.org","@type":"WebPage","@id":"https://daily.dev/posts/qwen3-8-max-alibaba-s-2-4-trillion-parameter-open-weight-model-benchmarked-0pqptx1df","comment":[{"@type":"Comment","text":"Even selling my two kidneys won’t be enough to run it locally, I guess!","datePublished":"2026-07-20T05:24:10.205Z","url":"https://daily.dev/posts/0pqptX1DF#c-wvCLFTb2c","author":{"@type":"Person","name":"Rahat Khan","url":"https://daily.dev/hidden_rahatkhan","image":"https://media.daily.dev/image/upload/s--RU21Sur4--/f_auto/v1774754359/avatars/avatar_b9ZHPMvBWPpXiOgr27Agq?_a=BAMAMiWQ0"},"interactionStatistic":{"@type":"InteractionCounter","interactionType":{"@type":"LikeAction"},"userInteractionCount":6}},{"@type":"Comment","text":"Beating Claude Opus 4 on benchmarks as an open weight model is wild but the note on inconsistent performance during long claude code sessions shows once again that multi turn agentic stability is the real benchmark developers actually care about","datePublished":"2026-07-20T09:58:15.776Z","url":"https://daily.dev/posts/0pqptX1DF#c-mcKNHDuoD","author":{"@type":"Person","name":"GeekLuffy","url":"https://daily.dev/geekluffy","image":"https://avatars.githubusercontent.com/u/71143694?v=4"},"interactionStatistic":{"@type":"InteractionCounter","interactionType":{"@type":"LikeAction"},"userInteractionCount":2}},{"@type":"Comment","text":"Closed models will always be superior - just logically. There’s no reason to open source something that’s superior. The second a developed model becomes superior, they’ll close it. Such is capitalism.","datePublished":"2026-07-20T21:25:54.074Z","url":"https://daily.dev/posts/0pqptX1DF#c-sYBNUO10N","author":{"@type":"Person","name":"Anthony Izzo","url":"https://daily.dev/amizzo","image":"https://lh3.googleusercontent.com/a/ACg8ocI3KpxrCJphZmqCBYWzrlZz0yUJS3pQF3RYarylAe8SP_G6_KUCoQ=s96-c"}},{"@type":"Comment","text":"What is seems to me with all those criticism over chinese models is that Americans keep to say it is not safe/legit/propaganda/propaganda/ etc etc to save a really big business for them. This is not safe, not an open weight model, baby!!!","datePublished":"2026-07-23T14:08:02.201Z","url":"https://daily.dev/posts/0pqptX1DF#c-Eiid88P8n","author":{"@type":"Person","name":"Simo's","url":"https://daily.dev/simos","image":"https://lh3.googleusercontent.com/a/AEdFTp7x7-eBPpw0hD5Wb7OiEwWMbKI3dMS62BANE0irnw=s96-c"}}]}
```

