> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.empiriolabs.ai/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.empiriolabs.ai/_mcp/server.

# Qwen3.5 397B-A17B

> Qwen3.5 397B-A17B is a flagship multimodal reasoning model for language, code, agents, GUI tasks, and image and video understanding.

![Qwen3.5 397B-A17B](https://media.empiriolabs.ai/model-logos/qwen.png)

[Alibaba Cloud](/providers/alibaba) · Text Generation

`POST /v1/chat/completions`

Qwen3.5 397B-A17B is a flagship multimodal reasoning model for language, code, agents, GUI tasks, and image and video understanding.

## At a glance

| Field               | Value                                                                                                                           |
| ------------------- | ------------------------------------------------------------------------------------------------------------------------------- |
| Model id            | `qwen3-5-397b-a17b`                                                                                                             |
| Model release date  | 2026-02-16                                                                                                                      |
| Input modalities    | Text, Image, Video                                                                                                              |
| Output modalities   | Text                                                                                                                            |
| Context window      | 256K                                                                                                                            |
| Weight precision    | -                                                                                                                               |
| Max output tokens   | 64,000                                                                                                                          |
| Region              | China                                                                                                                           |
| Features            | reasoning, vision, web\_search, function\_calling, multimodal                                                                   |
| Native inference    | No                                                                                                                              |
| New                 | No                                                                                                                              |
| Structured output   | JSON Mode                                                                                                                       |
| Supported endpoints | `POST /v1/chat/completions`, `POST /v1/responses`, `POST /v1/messages`, `POST /v1beta/models/qwen3-5-397b-a17b:generateContent` |
| Alternate model ids | `qwen3.5-397b-a17b`                                                                                                             |

## Pricing

| Charge     | Spec                    | Rate                                                        |
| ---------- | ----------------------- | ----------------------------------------------------------- |
| Input      | per 1M prompt tokens    | \<=128K \$0.172 (was \$0.60); 128K-256K \$0.43 (was \$0.60) |
| Output     | per 1M generated tokens | \<=128K \$1.032 (was \$3.60); 128K-256K \$2.58 (was \$3.60) |
| Web search | per call when invoked   | \$0.01                                                      |

## Example request

```bash
curl https://api.empiriolabs.ai/v1/chat/completions \
  -H 'Authorization: Bearer $EMPIRIOLABS_API_KEY' \
  -H 'Content-Type: application/json' \
  -d '{"model": "qwen3-5-397b-a17b", "messages": [{"role":"user","content":"Hello"}]}'
```

## Parameters

| Parameter                   | Type    | Required | Default    | Description                                                                                                                                                                            |
| --------------------------- | ------- | -------- | ---------- | -------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| `temperature`               | number  | no       | `0.7`      | Sampling temperature. 0 is deterministic and 2 is maximum randomness. · Range: 0 – 2                                                                                                   |
| `top_p`                     | number  | no       | `0.9`      | Nucleus sampling probability mass. Lower values make outputs more focused. · Range: 0 – 1                                                                                              |
| `max_tokens`                | number  | no       | `4096`     | Maximum output tokens. · Range: 1 – 64000                                                                                                                                              |
| `stop`                      | string  | no       | -          | Up to 4 strings where the model will stop generating further tokens.                                                                                                                   |
| `enable_thinking`           | boolean | no       | true       | Enable reasoning before answering.                                                                                                                                                     |
| `reasoning_effort`          | enum    | no       | `"medium"` | Reasoning effort level. none disables thinking. low, medium, high, and max set bounded thinking budgets sized to the selected model. · Allowed: `none`, `low`, `medium`, `high`, `max` |
| `thinking_budget`           | number  | no       | `32768`    | Maximum tokens reserved for reasoning when thinking is enabled. · Range: 1 – 80000                                                                                                     |
| `vl_high_resolution_images` | boolean | no       | true       | Use higher resolution processing for image inputs.                                                                                                                                     |
| `max_pixels`                | number  | no       | `2621440`  | Maximum pixel count per image when high resolution processing is disabled. · Range: 4096 – 16777216                                                                                    |
| `video_fps`                 | number  | no       | `2`        | Frames per second to sample from video inputs. · Range: 0.1 – 10                                                                                                                       |
| `tool_web_search`           | boolean | no       | false      | Search the web for real-time information. Adds \$0.01 to the request cost for each invoked call.                                                                                       |
| `response_format`           | enum    | no       | -          | Return the output as a valid JSON object (JSON mode). Describe the fields you want in your prompt.                                                                                     |

## Notes

Supports text, image, and video input. Web search is available through tool\_web\_search and adds \$0.01 per request when enabled. Thinking tokens are billed as output tokens.

---

*Machine-readable schema:* `GET https://api.empiriolabs.ai/v1/models/qwen3-5-397b-a17b`.