GLM-5.2

Z.ai (Zhipu AI) · zhipu/glm-5-2

ProviderZ.ai (Zhipu AI)
Familyglm
Typellm-reasoning
Statusactive
Released2026-06-13
Updated2026-09-10
Parameters753.3B
Open weightsyes

Capabilities

Chain Of Thought · tier-2 Function Calling · tier-2

What it fits on

Ordered by predicted speed. Bandwidth sets decode rate; memory decides whether it runs at all.

Every figure here is computed, not measured. Fit is weights at each quantisation against device memory, with a 25% allowance for the KV cache, activations and the OS. Decode rate is the memory-bandwidth roofline at 70% efficiency. Nobody has run this model on these devices.
DeviceMemoryBandwidthBest qualityWeightstok/sFastest
Apple M3 Ultra512.0 GB819.0 GB/sq4376.66 GB~1.5~1.5 q4

Where it runs

Taken from the card's availability section.

PlatformId on that platform
Hugging Facezai-org/GLM-5.2

Reported benchmark scores

This card reports no benchmark scores yet.

BenchmarkCatalogue standingScore

Data

This card as JSON · See it in the graph · Edit on GitHub