Qwen3.8 Max

POST /v1/chat/completionsTrillion-scale MoE flagship for coding, long-horizon agents, and professional work, with image and video understanding across a 1M-token context.
At a glance
Pricing
Example request
Parameters
Notes
Text, image, and video input are supported. Web search, web extractor, code interpreter, text-to-image search, and image-to-image search are optional built-in tools exposed through tool_* parameters. Web search, text-to-image search, and image-to-image search add $0.02 for each invoked call; web extractor and code interpreter run at no extra cost. Web extractor requires web search, and both web extractor and code interpreter require thinking. A single request can invoke a tool more than once, and each invoked call is billed. Thinking tokens are billed as output tokens. The qwen3-8-max id always serves the current Qwen3.8 Max build and moves with Alibaba’s snapshot upgrades; since September 5, 2026 it serves the qwen3-8-max-0902 snapshot. Use qwen3-8-max-0902 when output must stay pinned to that exact snapshot.
Per-tool billing (usage.tool_usage)
When this model invokes built-in tools inside a single request, the response carries a normalized usage.tool_usage map alongside the token counts:
Tool counts are already factored into cost_usd and are surfaced for transparency so you can audit per-tool billing. The field is omitted when no tools were invoked.
Variants
:variant1
Pricing
Parameters
Notes
Text, image, and video input are supported. Web search, web extractor, code interpreter, text-to-image search, and image-to-image search are optional built-in tools exposed through tool_* parameters. Web search, text-to-image search, and image-to-image search add $0.02 for each invoked call; web extractor and code interpreter run at no extra cost. Web extractor requires web search, and both web extractor and code interpreter require thinking. A single request can invoke a tool more than once, and each invoked call is billed. Thinking tokens are billed as output tokens. This id always serves the current Qwen3.8 Max build on the China endpoint and moves with Alibaba’s snapshot upgrades. Use qwen3-8-max-0902 when output must stay pinned to the September 2, 2026 snapshot.
Per-tool billing (usage.tool_usage)
When this model invokes built-in tools inside a single request, the response carries a normalized usage.tool_usage map alongside the token counts:
Tool counts are already factored into cost_usd and are surfaced for transparency so you can audit per-tool billing. The field is omitted when no tools were invoked.
Machine-readable schema: GET https://api.empiriolabs.ai/v1/models/qwen3-8-max.
