Log in

Qwen 3.8 Max

Default
Alibaba’s August 2026 frontier Max model on Model Studio: native multimodal (text, image, video) with a 1M-token context, up to 65K output, and hybrid reasoning via enable_thinking plus reasoning_effort (medium / xhigh). Built for agentic coding, tool use, and long-horizon multimodal workflows, with preserve_thinking for stable multi-turn agent context.
Alibaba’s August 2026 frontier Max model on Model Studio: native multimodal (text, image, video) with a 1M-token context, up to 65K output, and hybrid reasoning via enable_thinking plus reasoning_effort (medium / xhigh). Built for agentic coding, tool use, and long-horizon multimodal workflows, with preserve_thinking for stable multi-turn agent context.
Intelligence
Exceptional
Speed
Fast
Price
$2.00$6.00
Input • Output
Input
Text, image, audio
Output
Text, image, audio

Qwen 3.8 Max is alibaba’s august 2026 frontier max model on model studio: native multimodal (text, image, video) with a 1m-token context, up to 65k output, and hybrid reasoning via enable_thinking plus reasoning_effort (medium / xhigh). built for agentic coding, tool use, and long-horizon multimodal workflows, with preserve_thinking for stable multi-turn agent context..

1,000,000 context window
65,536 max output tokens
Mid 2026 knowledge cutoff
Reasoning token support
Pricing
Pricing is based on the number of tokens used. For tool-specific models, like search and computer use, there's a fee per tool call. See details in the pricing page.
Text tokens
Per 1M tokens
Batch API price
Input
$2.00
Cached input
$0.25
Output
$6.00
Quick comparison
Input
Cached input
Output
Qwen 3.8 Max
$2.00
Modalities
Text
Input and output
Image
Input only
Audio
Not supported
Endpoints
Chat Completions
v1/chat/completions
Responses
v1/responses
Batch
v1/batch
Realtime
v1/realtime
Assistants
v1/assistants
Fine-tuning
v1/fine-tuning
Embeddings
v1/embeddings
Image Generation
v1/images/generations
Image Edit
v1/images/edits
Speech Generation
v1/audio/speech
Transcription
v1/audio/transcriptions
Translation
v1/translations
Moderation
v1/moderations
Completions (legacy)
v1/completions
Features
Streaming
Supported
Function calling
Supported
Structured outputs
Not supported
Fine-tuning
Not supported
Distillation
Not supported
Fast response
Not supported
Cost efficient
Not supported
Tools
Tools supported by this model when using the Responses API.
Web search
Not supported
Code interpreter
Not supported
File search
Not supported
Image generation
Not supported
MCP
Not supported
Snapshots
Snapshots let you lock in a specific version of the model so that performance and behavior remain consistent. Below is a list of all available snapshots and aliases for Qwen 3.8 Max.
qwen3.8-max
Rate limits
Rate limits ensure fair and reliable access to the API by placing specific caps on requests or tokens used within a given time period. Your usage tier determines how high these limits are set and automatically increases as you send more requests and spend more on the API.
TierRPMTPMBatch queue limit
FreeNot supported
Tier 110010,000
Tier 2500100,000
Tier 310001,000,000
Tier 420002,000,000
Tier 550005,000,000