Qwen3 Max Model Specs, Costs & Benchmarks (July 2026)
Qwen3 Max, developed by Qwen, features a context window of 262.1K tokens. The model costs $0.78 per million tokens for input and $3.90 per million tokens for output. It was released on September 23, 2025, and has achieved impressive scores in various benchmarks.
Access Qwen3 Max & 210+ other AI models all in one platformTry Magica for free
Overview
Model Provider The organization behind this AI's development | |
Input Context Window Maximum input tokens this model can process at once | 262.1K tokens |
Output Token Limit Maximum output tokens this model can generate at once | 32.8K tokens |
Release Date When this model first became publicly available | September 23, 2025 10 months ago September 23rd, 2025 |
Knowledge Cutoff Latest training-data date reported by the provider | June 30, 2025 |
Pricing
Short Context Pricing Base rates before a long-context threshold applies | |
|---|---|
Input Token Cost Cost per million input tokens | $0.78 per million tokens |
Output Token Cost Cost per million output tokens | $3.90 per million tokens |
Cache Read Cost Cost to reuse cached input tokens | $0.16 per million tokens |
Cache Write Cost Cost to write input tokens to the cache | $0.97 per million tokens |
Long Context Pricing Rates for prompts with 32,000 or more tokens | |
Input Token Cost Long-context input cost per million tokens | $1.56 per million tokens |
Output Token Cost Long-context output cost per million tokens | $7.80 per million tokens |
Cache Read Cost Long-context cache read cost per million tokens | $0.31 per million tokens |
Cache Write Cost Long-context cache write cost per million tokens | $1.95 per million tokens |
Long Context Pricing Rates for prompts with 128,000 or more tokens | |
Input Token Cost Long-context input cost per million tokens | $1.95 per million tokens |
Output Token Cost Long-context output cost per million tokens | $9.75 per million tokens |
Cache Read Cost Long-context cache read cost per million tokens | $0.39 per million tokens |
Cache Write Cost Long-context cache write cost per million tokens | $2.44 per million tokens |
Capabilities & Features
Input Types Supported input formats | text |
Output Types Supported output formats | text |
Tokenizer Text encoding system | Qwen3 |
Key Features Advanced capabilities | β Function Callingβ Structured Output |
Reasoning Controls Defaults and configurable reasoning effort | Optional |
Benchmarks
Best Design Arena Score Highest human-preference Elo score across design arenas | 1,166 Elo (#38 in Asciiart (47.2% win rate)) |
Compare This Model
See how Qwen3 Max compares with other top models
Compare Qwen3 Max with top models in each category:
πProgramming
Best models for coding and development
π¨Creative & Roleplay
Models optimized for creative writing
π’Marketing
Content creation and marketing tasks
π»Technology
Technical analysis and explanations
π¬Science
Scientific research and analysis
πTranslation
Multilingual translation tasks