Context
1M
Input
text
Output
text
Tool calling
Supported
About this model
GLM-5.3 is Z.ai's advanced Mixture-of-Experts model, building upon the GLM-5.2 base architecture with significantly scaled post-training. It is designed specifically for complex coding, long-horizon tasks, and autonomous agent workflows. The model sets a new state-of-the-art for open models in agentic coding and demonstrates strong emergent cybersecurity capabilities, excelling in vulnerability discovery and exploitation chaining.
Input modalities
text
Accepted as model input
Providers
Available routing options for this model through OneInfer.
Available
Input tokens
$1.400 / 1M tokens
Output tokens
$4.400 / 1M tokens
Routing
OneInfer optimized
Pricing
Current OneInfer pricing for this model.
| Usage | Price |
|---|---|
| Input tokens | $1.400 / 1M tokens |
| Output tokens | $4.400 / 1M tokens |
| Cached input tokens | $0.260 / 1M tokens |
Performance
Published evaluation results associated with this model.
Coding
Terminal Bench 2.188.2
Terminal Bench 3.028.3
DeepSWE v1.166.9
NL2Repo58
FrontierSWE78.1
SWE-Marathon v1.142.5
PostTrainBench39.8
Cyber
CyberGym84.5
ExploitGym105
ExploitBench54.4
Agentic
Toolathlon Verified73
AutomationBench48.2
Agents' Last Exam28.5
HLE w/ Tools62.5
GDPval-AA v21769
API example
curl https://api.oneinfer.ai/v1/ula/chat/completions \
-H "Authorization: Bearer $JWT_TOKEN" \
-H "Content-Type: application/json" \
-d '{
"model": "zai-org/GLM-5.1",
"messages": [
{ "role": "user", "content": "Hello!" }
]
}'