GLM-4.5

glm-4.5

GLM-4.5 is an agent-oriented MoE model from Z.ai, described with 355B total parameters and 32B activated parameters. Official docs highlight reasoning, coding, tool use, and browser-style agent abilities, with both thinking and non-thinking modes. It works well as the GLM line’s earlier agent foundation model before GLM-5.

Context Window

131.1K tokens

Maximum Output

98.3K tokens

Release Date

Jul 28, 2025

Modalities

GLM-4.5 Pricing

Input PriceOutput Price
$0.4286/M$2/M

GLM-4.5 API Capabilities

Reasoning

Supported

Tool calling

Supported

Temperature parameter

Supported

Attachments

Not supported

Knowledge Base

2025-04-01

Endpoint Protocols

geminiCompletions APIMessages API

GLM-4.5 Model Highlights

GLM-4.5 is an open-weight text model that unifies reasoning, coding and agent capabilities with hybrid thinking modes and a 128K context.

Unified Agentic Reasoning

Combines reasoning, coding and agent-oriented training in one model for multi-step problem solving and software tasks.

Hybrid Thinking Modes

Offers a thinking mode for complex reasoning and an immediate-response mode for simpler workloads, allowing developers to match effort to the task.

Open-Weight MoE

Uses a mixture-of-experts architecture with 355B total and 32B active parameters, released under the MIT license for independent deployment.

GLM-4.5 Use Cases

GLM-4.5 suits software engineering agents, complex text reasoning and self-hosted development workflows that benefit from selectable thinking depth.

Software Engineering Agent

Plan code changes, inspect project files and carry out implementation and debugging steps across a multi-turn development task.

Complex Text Reasoning

Analyze long technical requirements or multi-step logic problems in thinking mode and produce a reasoned textual solution.

Private Model Deployment

Deploy the MIT-licensed open weights in controlled infrastructure for internal coding, reasoning or agent workloads.

How to Use GLM-4.5 via the TokenHub API

Create API key

Replace these path values before running: {model}

curl 'https://us-api.tokenhub.com/v1beta/models/{model}:generateContent' \
  -X 'POST' \
  -H "Authorization: Bearer $TOKENHUB_API_KEY"

GLM-4.5 Benchmarks

GLM-4.5 (Reasoning)

Index score
Artificial Analysis Intelligence IndexArtificial Analysis broad capability aggregate19.5
Artificial Analysis Coding IndexArtificial Analysis software task aggregate26.3
Artificial Analysis Math IndexArtificial Analysis math reasoning aggregate73.7
Knowledge & Reasoning
MMLU-ProAdvanced multi-task knowledge83.5%
GPQAAdvanced science problem solving78.2%
HLEBroad expert-level exam set12.2%
Coding & Engineering
LiveCodeBenchLive coding problems73.8%
SciCodeScientific coding challenges34.8%
Terminal-Bench HardHard terminal task execution22.0%
Math
MATH-500Advanced math problem solving97.9%
AIMECompetition math problems87.3%
AIME 2025Competition math problems73.7%
Instruction Following & Agent Tasks
IFBenchPrompt constraint adherence44.1%
AA-LCRLong-context reasoning48.3%
τ²-BenchAgent workflow tasks43.0%

Metrics sourced from Artificial Analysis

Media and Discussions

Selected public videos and posts related to this model.

X (Twitter)

View post on X
View post on X
View post on X

Reddit

YouTube

Watch on YouTube
Watch on YouTube
Watch on YouTube

GLM-4.5 FAQ

GLM-4.5: capabilities, use cases, limits, and TokenHub guidance.

What is GLM-4.5?+

GLM-4.5 is a Z.AI model for reasoning, coding, and native agent workflows.

Which workloads suit GLM-4.5?+

Best for code reasoning, agent workflows and tool-heavy automation, especially when deep reasoning is the priority.

Which feature stands out?+

Key strength: a unified focus on reasoning, coding, and native agents.

When should teams avoid GLM-4.5?+

It belongs to an older generation and may lack newer capabilities. For the latest capabilities matter, consider GLM-5.

What should I verify in TokenHub?+

Confirm TokenHub availability; prefer the current successor for new work.

Ready to use GLM-4.5?

Use one API key to access GLM-4.5 and more AI models through TokenHub.

Create API key