Models

Coding GLM 5 vs Coding GLM 5 (free)

Compare Coding GLM 5 from Z.AI and Coding GLM 5 (free) from Z.AI on key metrics including benchmarks, price, context length, and other model features. Access both models and hundreds of others through the AIHubMix API.

Z.AICoding GLM 5Z.AICoding GLM 5 (free)
Z.AI logo
Coding GLM 5
Z.AI · text → text

Only supports OpenAI-compatible formats.

Input$0.06 /M
Output$0.22 /M
Z.AI logo
Coding GLM 5 (free)
Z.AI · text → text

coding-glm-5-free is the open and free version of coding-glm-5. To ensure stable service performance, usage limits are in place: up to 5 requests per minute, 100 requests per day, and a daily token allowance of 1 million.

Input$0 /M
Output$0 /M

Pricing & Specifications

Prices are per million tokens. Time to First Token and throughput are rolling averages measured on AIHubMix.

Coding GLM 5
Coding GLM 5 (free)
Input /M
$0.06
$0
Output /M
$0.22
$0
Cache read /M
-
-
Context length
200,000
200,000
Max output
131,072
131,072
Time to First Token
1.7 s
6.2 s
Throughput
50.3 tok/s
32.0 tok/s
Modalities
text
text
Supported Parameters
thinkingtoolsfunction callingstructured outputs
thinkingtoolsfunction callingstructured outputs
API Formats
Released
February 12, 2026
February 11, 2026

Promotional prices show the discounted rate; see each model page for promotion windows.

Activity Past 30 Days

Daily traffic served through AIHubMix — how demand for each model is trending.

coding-glm-5coding-glm-5-free

Tokens / day

-

Requests / day

-

Performance Past 3 Days

Measured on real AIHubMix traffic, hourly buckets. Gaps mean no traffic in that hour.

coding-glm-5coding-glm-5-free

Throughput (tok/s)

-

TTFT (s)

-

Uptime (%)

-

LMArena Benchmarks

LMArena ratings by capability (Bradley-Terry, commonly called Elo). Higher is better.

Text
coding-glm-5coding-glm-5-free
1420146015001540
Overall
14571458
Coding
14961498
Math
14391443
Hard prompts
14781479
Instruction following
14471448
Multi-turn
14711473
Creative writing
14461450
Longer query
14701471
Chinese
15241526
English
14711471
WebDev Arena
coding-glm-5coding-glm-5-free
14001420144014601480
Overall
14301435
React
14211425
HTML
14431452
Gaming
14371437
Simulations
14411441
Data analytics
14291429

Source: LMArena (arena.ai) leaderboard, imported by AIHubMix. Models without published ratings are omitted per chart.

Cost calculator

Estimate your monthly bill for the same workload on each model.

Coding GLM 5 (free)
$0.00 /mo
Coding GLM 5
$6.90 /mo

Monthly = daily × 30. Discounted rates applied where a promotion is active.

FAQ

How do their coding arena scores compare?

Coding GLM 5: 1498; Coding GLM 5 (free): 1496 (LMArena coding leaderboard).

Which responds faster?

Coding GLM 5: 1.7s time to first token measured on AIHubMix; see the live performance charts above for how each model behaves across the day.

How large is each context window?

Coding GLM 5 accepts 200,000 and Coding GLM 5 (free) accepts 200,000 input tokens. Maximum output per request is 131,072 tokens on Coding GLM 5 and 131,072 tokens on Coding GLM 5 (free).

Which one generates tokens faster?

Coding GLM 5 at 50.3 tok/s and Coding GLM 5 (free) at 32.0 tok/s, measured as output throughput on AIHubMix — a separate metric from time to first token.

What inputs and capabilities does each model support?

Coding GLM 5 accepts text input and supports thinking, tool calling, function calling and structured outputs; Coding GLM 5 (free) accepts text input and supports thinking, tool calling, function calling and structured outputs.

Can I call Coding GLM 5 and Coding GLM 5 (free) with the same API key?

Yes. AIHubMix serves every model on this page behind one OpenAI-compatible endpoint, so switching between them is a one-line change to the model field — no second account, key or SDK.

Popular comparisons

Related model match-ups readers also look at.