> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.empiriolabs.ai/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.empiriolabs.ai/_mcp/server.

# GLM 5.3

> Coding and agentic model with a 1M token context, 128K output, always-on reasoning at low, high, or max effort, native web search, and tool calling.

![GLM 5.3](https://media.empiriolabs.ai/model-logos/glm.png)

[Z.ai](/providers/zhipu) · Text Generation

`POST /v1/chat/completions`

Coding and agentic model with a 1M token context, 128K output, always-on reasoning at low, high, or max effort, native web search, and tool calling.

## At a glance

| Field               | Value                                                                                                                 |
| ------------------- | --------------------------------------------------------------------------------------------------------------------- |
| Model id            | `glm-5-3`                                                                                                             |
| Model release date  | 2026-08-14                                                                                                            |
| Input modalities    | Text                                                                                                                  |
| Output modalities   | Text                                                                                                                  |
| Context window      | 1M                                                                                                                    |
| Weight precision    | -                                                                                                                     |
| Max output tokens   | 131,072                                                                                                               |
| Region              | Singapore                                                                                                             |
| Features            | reasoning, function\_calling, web\_search                                                                             |
| Native inference    | No                                                                                                                    |
| New                 | Yes                                                                                                                   |
| Structured output   | JSON Mode                                                                                                             |
| Supported endpoints | `POST /v1/chat/completions`, `POST /v1/responses`, `POST /v1/messages`, `POST /v1beta/models/glm-5-3:generateContent` |
| Alternate model ids | `glm-5.3`, `zai/glm-5.3`, `zhipu/glm-5.3`                                                                             |

## Pricing

| Charge     | Spec                    | Rate    |
| ---------- | ----------------------- | ------- |
| Input      | per 1M prompt tokens    | \$1.40  |
| Output     | per 1M generated tokens | \$4.40  |
| Web search | per request             | \$0.033 |

## Example request

```bash
curl https://api.empiriolabs.ai/v1/chat/completions \
  -H 'Authorization: Bearer $EMPIRIOLABS_API_KEY' \
  -H 'Content-Type: application/json' \
  -d '{"model": "glm-5-3", "messages": [{"role":"user","content":"Hello"}]}'
```

## Parameters

| Parameter               | Type    | Required | Default     | Description                                                                                                                                                                                                           |
| ----------------------- | ------- | -------- | ----------- | --------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| `max_tokens`            | integer | no       | `65536`     | Maximum number of output tokens to generate. · Range: 1 – 131072                                                                                                                                                      |
| `temperature`           | number  | no       | `1`         | Controls randomness. Lower values make responses more deterministic. · Range: 0 – 1                                                                                                                                   |
| `top_p`                 | number  | no       | `0.95`      | Nucleus sampling cutoff. · Range: 0.01 – 1                                                                                                                                                                            |
| `reasoning_effort`      | enum    | no       | `"max"`     | GLM 5.3 reasoning effort. This model always reasons and cannot be turned off. low is lightweight, high is enhanced, and max is deep reasoning. max is recommended for complex coding. · Allowed: `low`, `high`, `max` |
| `do_sample`             | boolean | no       | true        | Enable sampling. Turn off for greedy deterministic output (temperature and top\_p are ignored).                                                                                                                       |
| `tool_web_search`       | boolean | no       | false       | Enable built-in web search. Adds \$0.033 per request when used.                                                                                                                                                       |
| `search_recency_filter` | enum    | no       | `"noLimit"` | Limit web search results to a recency window. · Allowed: `oneDay`, `oneWeek`, `oneMonth`, `oneYear`, `noLimit`                                                                                                        |
| `count`                 | integer | no       | `10`        | Number of web search results to retrieve when web search is enabled. · Range: 1 – 50                                                                                                                                  |
| `search_domain_filter`  | string  | no       | -           | Restrict web search to a specific domain.                                                                                                                                                                             |
| `search_prompt`         | string  | no       | -           | Optional prompt used to summarize retrieved web search results.                                                                                                                                                       |
| `search_result`         | boolean | no       | true        | Return web search result metadata in the response when web search is enabled.                                                                                                                                         |
| `tool_stream`           | boolean | no       | false       | Stream function-call arguments incrementally when streaming.                                                                                                                                                          |
| `tools`                 | array   | no       | `[]`        | OpenAI-compatible function calling tool definitions.                                                                                                                                                                  |
| `tool_choice`           | object  | no       | -           | OpenAI-compatible tool choice control.                                                                                                                                                                                |
| `stop`                  | array   | no       | -           | Optional stop sequences (up to 4).                                                                                                                                                                                    |
| `response_format`       | enum    | no       | -           | Return the output as a valid JSON object (JSON mode). Describe the fields you want in your prompt.                                                                                                                    |

---

*Machine-readable schema:* `GET https://api.empiriolabs.ai/v1/models/glm-5-3`.