> For clean Markdown of any page, append .md to the page URL. > For a complete documentation index, see https://docs.empiriolabs.ai/models/glm-5-3/llms.txt. > For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.empiriolabs.ai/_mcp/server. # GLM 5.3 > Coding and agentic model with a 1M token context, 128K output, always-on reasoning at low, high, or max effort, native web search, and tool calling. ![GLM 5.3](https://media.empiriolabs.ai/model-logos/glm.png) [Z.ai](/providers/zhipu) · Text Generation `POST /v1/chat/completions` Coding and agentic model with a 1M token context, 128K output, always-on reasoning at low, high, or max effort, native web search, and tool calling. ## At a glance | Field | Value | | ------------------- | --------------------------------------------------------------------------------------------------------------------- | | Model id | `glm-5-3` | | Model release date | 2026-08-14 | | Input modalities | Text | | Output modalities | Text | | Context window | 1M | | Weight precision | - | | Max output tokens | 131,072 | | Region | Singapore | | Features | reasoning, function\_calling, web\_search | | Native inference | No | | New | Yes | | Structured output | JSON Mode | | Supported endpoints | `POST /v1/chat/completions`, `POST /v1/responses`, `POST /v1/messages`, `POST /v1beta/models/glm-5-3:generateContent` | | Alternate model ids | `glm-5.3`, `zai/glm-5.3`, `zhipu/glm-5.3` | ## Pricing | Charge | Spec | Rate | | ---------- | ----------------------- | ------- | | Input | per 1M prompt tokens | \$1.40 | | Output | per 1M generated tokens | \$4.40 | | Web search | per request | \$0.033 | ## Example request ```bash curl https://api.empiriolabs.ai/v1/chat/completions \ -H 'Authorization: Bearer $EMPIRIOLABS_API_KEY' \ -H 'Content-Type: application/json' \ -d '{"model": "glm-5-3", "messages": [{"role":"user","content":"Hello"}]}' ``` ## Parameters | Parameter | Type | Required | Default | Description | | ----------------------- | ------- | -------- | ----------- | --------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | | `max_tokens` | integer | no | `65536` | Maximum number of output tokens to generate. · Range: 1 – 131072 | | `temperature` | number | no | `1` | Controls randomness. Lower values make responses more deterministic. · Range: 0 – 1 | | `top_p` | number | no | `0.95` | Nucleus sampling cutoff. · Range: 0.01 – 1 | | `reasoning_effort` | enum | no | `"max"` | GLM 5.3 reasoning effort. This model always reasons and cannot be turned off. low is lightweight, high is enhanced, and max is deep reasoning. max is recommended for complex coding. · Allowed: `low`, `high`, `max` | | `do_sample` | boolean | no | true | Enable sampling. Turn off for greedy deterministic output (temperature and top\_p are ignored). | | `tool_web_search` | boolean | no | false | Enable built-in web search. Adds \$0.033 per request when used. | | `search_recency_filter` | enum | no | `"noLimit"` | Limit web search results to a recency window. · Allowed: `oneDay`, `oneWeek`, `oneMonth`, `oneYear`, `noLimit` | | `count` | integer | no | `10` | Number of web search results to retrieve when web search is enabled. · Range: 1 – 50 | | `search_domain_filter` | string | no | - | Restrict web search to a specific domain. | | `search_prompt` | string | no | - | Optional prompt used to summarize retrieved web search results. | | `search_result` | boolean | no | true | Return web search result metadata in the response when web search is enabled. | | `tool_stream` | boolean | no | false | Stream function-call arguments incrementally when streaming. | | `tools` | array | no | `[]` | OpenAI-compatible function calling tool definitions. | | `tool_choice` | object | no | - | OpenAI-compatible tool choice control. | | `stop` | array | no | - | Optional stop sequences (up to 4). | | `response_format` | enum | no | - | Return the output as a valid JSON object (JSON mode). Describe the fields you want in your prompt. | --- *Machine-readable schema:* `GET https://api.empiriolabs.ai/v1/models/glm-5-3`. > Coding and agentic model with a 1M token context, 128K output, always-on reasoning at low, high, or max effort, native web search, and tool calling.