Qwen3.8 Max 0902

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text, with a 1M-token context window and reasoning enabled by default.

This snapshot is post-trained for coding and agentic work, including multi-step software projects, multi-tool orchestration, and long-horizon task execution. It also targets chart reasoning, document parsing, and multimodal understanding over long documents and extended video. Tool calling, structured outputs, and configurable reasoning effort are supported.

alibaba/qwen3.8-max-0902
Organization
Alibaba
Family
qwen
Providers
6
Context
1,000,000
Output limit
131,072
Knowledge
—
Release
2026-09-02
Updated
2026-09-02
Weights
Closed
Input
text, image, video, pdf
Output
text
Capabilities
Tools, Reasoning, Temperature

Providers

Compare every cataloged offer for this model, including its provider, route, token limits, price, and capabilities.

7 offers · USD per million tokens
ProviderModel IDContextOutputInput / output · 1MReasoningToolsStructuredDetails
Eden AI
qwen/qwen3.8-max-0902
1,000,000131,072$2.00 / $6.00YesYesYes
View
Kilo Gateway
qwen/qwen3.8-max-0902
1,000,000131,072$2.00 / $6.00YesYesYes
View
NanoGPT
qwen/qwen3.8-max-0902
991,808131,072$2.00 / $6.00YesYesYes
View
Ofox
bailian/qwen3.8-max-0902
1,000,000131,072$2.00 / $6.00YesYes—
View
Ofox
qwen/qwen3.8-max-0902
1,000,000131,072$1.71 / $5.14YesYes—
View
Alibabavia OpenRouter
qwen/qwen3.8-max-0902
unknown
1,000,000131,072$2.00 / $6.00YesYesYes
View
Vercel AI Gateway
alibaba/qwen3.8-max-0902
991,000128,000$2.00 / $6.00YesYes—
View

Prices and limits apply to each offer. A dash means the value is not provided. OpenRouter offers show the hosting provider and routing channel separately.

Pricing

Compare token rates and cache charges across available routes and hosting configurations.

USD per million tokens
Input from$1.71Per 1M input tokens
Output from$5.14Per 1M output tokens
Available offers7Across 6 providers

Lowest listed input and output prices across the offers above; they may belong to different providers. These are base rates, before conditional pricing, discounts or additional fees.

Endpoint price comparisonLatest observed snapshot · 1 lowest-priced configuration
InputOutput
Alibabaalibaba
$2.00$6.00
qwen/qwen3.8-max-0902Observed
Input
$2.00
Output
$6.00
Cache read
$0.25
Cache write
$2.50
Hosting prices through OpenRouter All endpoints and cache rates
EndpointInput / 1MOutput / 1MCache read / 1MCache write / 1M
Alibabaalibaba$2.00$6.00$0.25$2.50

Each row is a hosting configuration. Open an offer in Providers for its full pricing conditions and additional fees. Endpoint observations may differ from the routing catalog's base price.

Performance

Review reported latency and output-throughput percentiles from the latest observed rolling window.

Latest 30-minute observation

Time to first token · milliseconds

Latency distributionRolling 30-minute snapshot · 1 endpoint
P50P75P90P99
Alibabaalibaba
4,140.5–49,607.83 ms
All latency percentiles 1 endpoints
EndpointP50P75P90P99
Alibabaalibaba4,140.5 ms7,018.25 ms11,542.8 ms49,607.83 ms

Output throughput · tokens per second

Throughput distributionRolling 30-minute snapshot · 1 endpoint
P50P75P90P99
Alibabaalibaba
34–58 t/s
All throughput percentiles 1 endpoints
EndpointP50P75P90P99
Alibabaalibaba34 t/s41 t/s47 t/s58 t/s

Percentiles describe the observed request distribution. These are snapshots of a rolling window, not a historical time series.

Uptime

Compare successful request rates across the latest 5-minute, 30-minute, and 24-hour observations.

Successful requests by window
Availability by endpointCurrent overlapping windows · 1 endpoint
Endpoint5 minutes30 minutes24 hours
Alibabaalibaba

Reported by OpenRouter at the time of collection; rate-limited requests are excluded. A dash means no measurement was reported. This is not a live availability check.

Benchmarks

Review attributed evaluation results on each source's original scale.

External evaluations
qwen/qwen3.8-max-0902Observed
Benchmark profileArtificial Analysis · shared source scale from 0 to 80
Coding76.2
Agentic56
Intelligence45.4

Sources: Artificial Analysis and Design Arena, via OpenRouter. Scores retain their original scales and are not OpenAI Suite ratings.

Detailed evaluations

EvaluationConfigurationSourceResultsDetails
Composite indicesQwen3.8 Max (0902)qwen/qwen3.8-max-20260902Artificial AnalysisIntelligence 45.4 · Coding 76.2 · Agentic 56
Details

Sources: Artificial Analysis, Design Arena and OpenRouter. Evaluation configurations and scales differ; results are not interchangeable. Updated Oct 4, 2026.

Apps & session costs

Compare observed median session costs for applications using this model, grouped by conversation length.

30-day sample · through Sep 27, 2026
ApplicationModel routeTurnsMedian cost / sessionDetails
Hermes Agentqwen/qwen3.8-max-2026090250+ turns
$3.334543
Details
Claude Codeqwen/qwen3.8-max-2026090250+ turns
$6.066278
Details
Hermes Agentqwen/qwen3.8-max-2026090210–49 turns
$0.466432
Details
Hermes Agentqwen/qwen3.8-max-202609021 turn
$0.003517
Details
Claude Codeqwen/qwen3.8-max-2026090210–49 turns
$1.070254
Details
Claude Codeqwen/qwen3.8-max-202609022–9 turns
$0.135476
Details
Hermes Agentqwen/qwen3.8-max-202609022–9 turns
$0.063501
Details

Source: OpenRouter, as of Oct 2, 2026. Published under CC BY 4.0. These are observed costs per session, not token prices or a forecast for your workload. Apps are compared separately.

Activity

Daily prompt and completion tokens reported for this model's routes in OpenRouter's top 50. Missing days have no published value.

Last 90 complete UTC days
Daily token usage952,844,469,829 reported tokens · 17 days with observations
Jul 6, 2026Oct 3, 2026
Daily observations · 17 days
Date (UTC)Reported tokensReported routes
Oct 3, 202646,677,448,6681
Sep 22, 202661,349,292,3371
Sep 20, 202637,682,735,3071
Sep 19, 202665,294,508,6961
Sep 18, 202688,309,969,5001
Sep 17, 202653,556,546,3611
Sep 16, 202653,192,607,1631
Sep 15, 202641,590,160,8681
Sep 14, 202647,041,645,7621
Sep 13, 202637,687,003,5681
Sep 12, 202663,710,750,3291
Sep 10, 202681,231,424,4901
Sep 9, 202668,375,093,4091
Sep 8, 202662,584,798,1801
Sep 7, 202670,379,635,9171
Sep 6, 202638,503,040,4081
Sep 5, 202635,677,808,8661

Source: OpenRouter (openrouter.ai/rankings), as of Oct 4, 2026. CC BY 4.0. Totals include only individually published routes. Tokenizers differ by provider; the aggregated “other” category is never assigned to a model.

Technical details & API

Explore specifications, supported parameters, reasoning controls, and API identifiers for each OpenRouter route.

Current route specifications
qwen/qwen3.8-max-0902Observed
Context
1,000,000
Maximum output
131,072
Tokenizer
Qwen
Input
text, image, video
Output
text
Moderation
No
Added to OpenRouter
2026-09-03
Base URLhttps://openrouter.ai/api/v1
Model IDqwen/qwen3.8-max-0902
Version IDqwen/qwen3.8-max-20260902

Reasoning

Required
Yes
Enabled by default
Yes
Default effort
xhigh
Supported efforts
xhigh, high, medium, low, minimal

Supported parameters

  • frequency_penalty
  • include_reasoning
  • logprobs
  • max_tokens
  • presence_penalty
  • reasoning
  • reasoning_effort
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_k
  • top_logprobs
  • top_p

Frequently asked questions

Answers based on the model metadata and serving offers currently in the catalog.

How much does Qwen3.8 Max 0902 cost?

Listed rates start at $1.71 per million input tokens and $5.14 per million output tokens. Prices vary by provider and configuration; see Pricing for details.

Which providers offer Qwen3.8 Max 0902?

There are 7 cataloged offers across 6 providers. Compare model IDs, prices and limits in Providers.

What is the context window?

The cataloged model context is 1,000,000 tokens. Each serving endpoint may apply a different limit.

Does it support tools and structured output?

Tool calling: Yes. Structured output: Not reported. Support can vary by endpoint.

Catalog history · 2 recorded revisions

Metadata observations since this model was first cataloged. These are not model release versions.

  • · d80589c75d21
  • · 2ae995e64bef