Model comparison

GLM-4.5-Air vs Sonar

GLM-4.5-Air and Sonar score almost the same on the Noometry Index (38.9 vs 38.5), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

GLM-4.5-Air Z.ai (Zhipu)

38.9

Rank #177 Confirmed

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in writing & preference, where GLM-4.5-Air leads 55.9 to 52.6.
  • GLM-4.5-Air is cheaper at $0.20 / $1.10 per million input/output tokens, against $1 / $1 for Sonar.
  • GLM-4.5-Air accepts more context: 131K tokens versus 128K.
  • GLM-4.5-Air has downloadable open weights; the other is API-only.

Side by side

GLM-4.5-Air and Sonar specifications
GLM-4.5-AirSonar
ProviderZ.ai (Zhipu)Perplexity
Noometry Index38.938.5
Released2025-07-202024-01-01
WeightsOpenProprietary
Context window131K128K
Max output98K4K
Input $ / M tokens$0.20$1
Output $ / M tokens$1.10$1
Results tracked277

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Sonar leads

GLM-4.5-Air: 33.3 (#259), Sonar: 35.7 (#221)

Coding benchmarks
BenchmarkGLM-4.5-AirSonar
GSO2.9%—
LiveBench Coding—35.1%
LMArena Coding1397—

Reasoning GLM-4.5-Air leads

GLM-4.5-Air: 24.1 (#166), Sonar: 21.1 (#227)

Reasoning benchmarks
BenchmarkGLM-4.5-AirSonar
Kagi LLM Benchmark43%—
LiveBench Reasoning—46.3%
LMArena Hard Prompts1379—
LiveBench Data Analysis—37.9%
ForecastBench59.2—
LiveBench—46.9%

Math GLM-4.5-Air leads

GLM-4.5-Air: 36.2 (#170), Sonar: 33.7 (#200)

Math benchmarks
BenchmarkGLM-4.5-AirSonar
Omni-MATH39.1%—
LiveBench Math—41.6%
LMArena Math1396—

Knowledge Not comparable

GLM-4.5-Air: 35.0 (#191), Sonar: —

Knowledge benchmarks
BenchmarkGLM-4.5-AirSonar
Humanity's Last Exam8.1%—
MMLU-Pro76.2%—
Vectara Hallucination Rate9.3%—
GPQA (HELM)59.4%—
LMArena Expert1370—

Multilingual Not comparable

GLM-4.5-Air: 49.1 (#135), Sonar: —

Multilingual benchmarks
BenchmarkGLM-4.5-AirSonar
LMArena Non-English1366—
LMArena Chinese1426—
LMArena French1399—
LMArena German1377—
LMArena Japanese1348—
LMArena Korean1308—
LMArena Russian1373—
LMArena Spanish1386—

Instruction Following Sonar leads

GLM-4.5-Air: 69.6 (#171), Sonar: 71.4 (#150)

Instruction Following benchmarks
BenchmarkGLM-4.5-AirSonar
LiveBench Instruction Following—76.2%
IFEval81.2%—
LMArena Instruction Following1354—

Long Context Not comparable

GLM-4.5-Air: 41.6 (#135), Sonar: —

Long Context benchmarks
BenchmarkGLM-4.5-AirSonar
LMArena Longer Query1366—

Writing & Preference GLM-4.5-Air leads

GLM-4.5-Air: 55.9 (#139), Sonar: 52.6 (#167)

Writing & Preference benchmarks
BenchmarkGLM-4.5-AirSonar
LMArena Text1384—
LMArena Creative Writing1343—
WildBench78.9%—
LMArena Multi-Turn1371—
LiveBench Language—44.1%

Frequently asked questions

Is GLM-4.5-Air better than Sonar?

GLM-4.5-Air and Sonar score almost the same on the Noometry Index (38.9 vs 38.5), so choose on price, context window or the category you care about most.

Which is cheaper, GLM-4.5-Air or Sonar?

GLM-4.5-Air is cheaper. It lists at $0.20 per million input tokens and $1.10 per million output tokens; Sonar lists at $1 and $1.

Is GLM-4.5-Air or Sonar better for coding?

Sonar scores higher on coding benchmarks: 35.7 versus 33.3 in the Noometry coding category.

Which has the bigger context window?

GLM-4.5-Air does, with 131K tokens against 128K.

How many benchmarks do GLM-4.5-Air and Sonar share?

0 benchmarks have published results for both models. GLM-4.5-Air has 27 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper