Zhipu: healthcare benchmark results
1 model · 3 results · snapshot reviewed September 28, 2026
Everything the index holds for Zhipu, gathered in one place: GLM 5.2. Scores sit on each benchmark's own scale and never compare across rows from different boards.
Every result
| model | benchmark | score | index position |
|---|---|---|---|
| GLM 5.2 Subject suites release set: 257 tasks across eight subjects; one answer per task, no tools; blind cross-family grading. 95% bootstrap CI 18.1–23.4. Harness snapshot September 10, 2026; not the separate 89-task incretin ranking. | Health Optimization Bench | 20.7 | 12 of 16 |
| GLM 5.2 model ID zai/glm-5.2; temperature=1; max_output_tokens=30000; standard error 2.166 pp; $0.045010/test; source snapshot 2026-09-26; run date not published | MedCode (Vals AI) | 40.77% | 56 of 102 |
| GLM 5.2 model ID zai/glm-5.2; temperature=1; max_output_tokens=30000; standard error 2.002 pp; $0.044912/test; source snapshot 2026-09-26; run date not published | MedScribe (Vals AI) | 83.53% | 42 of 104 |
Positions refer to indexed rows, including configuration variants, and are not controlled comparisons across sources or graders. Source board sizes appear separately where coverage differs.
Which healthcare benchmarks does Zhipu appear on?
As of September 28, 2026, Zhipu models hold 3 indexed results across 3 tracked benchmarks, through GLM 5.2.
Where does Zhipu have the highest indexed score?
Zhipu does not have the highest indexed score on any tracked board in this snapshot.
Other labs with pages: OpenAI, Anthropic, Google, Alibaba, Meta, Moonshot AI, DeepSeek, SpaceXAI, xAI, Xiaomi, MiniMax, Thinking Machines, zAI, NVIDIA, SpaceX AI, Zhipu AI, Mistral, Ant Group, Poolside, Z.ai, Microsoft, Baichuan, Tencent, Inception, Cohere. The full field is on the index.