Models / GLM 4.7 Flash

Z.ai

GLM 4.7 Flash

15 published results from 3 sources. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.

Provider
Z.ai
Sources
3
Our benchmarks
0
Price
Not yet published

Reported by others

1,370
Overall · rank 157 of 177
Unit
Arena rating, higher is better
Range
1,365 to 1,376
Sample
11855 votes
Configuration
GLM 4.7 Flash
Measured
25 Sep 2026
Not shown
Not an error rate: claims that cannot be checked on the web are skipped, and preference still carries most of the weight.
1,368
Business, management and finance · rank 165 of 402
Unit
Arena rating, higher is better
Range
1,356 to 1,380
Sample
2244 votes
Configuration
GLM 4.7 Flash
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,306
Creative writing · rank 191 of 407
Unit
Arena rating, higher is better
Range
1,293 to 1,320
Sample
1871 votes
Configuration
GLM 4.7 Flash
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,386
Expert prompts · rank 146 of 359
Unit
Arena rating, higher is better
Range
1,367 to 1,406
Sample
885 votes
Configuration
GLM 4.7 Flash
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,348
Instruction following · rank 185 of 409
Unit
Arena rating, higher is better
Range
1,338 to 1,358
Sample
3298 votes
Configuration
GLM 4.7 Flash
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,365
Overall · rank 185 of 409
Unit
Arena rating, higher is better
Range
1,359 to 1,371
Sample
11983 votes
Configuration
GLM 4.7 Flash
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,331
Writing, literature and language · rank 187 of 408
Unit
Arena rating, higher is better
Range
1,320 to 1,343
Sample
2657 votes
Configuration
GLM 4.7 Flash
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
Reported by UGI Leaderboard
172.0%
Requested-length error · rank 364 of 370
Unit
% off the requested word count, lower is better
Configuration
glm-4.7-flash
Measured
3 Feb 2026
Not shown
Not other format limits such as character counts or bullet counts.
Reported by UGI Leaderboard
39.0%
Requested-length error · rank 290 of 370
Unit
% off the requested word count, lower is better
Configuration
glm-4.7-flash
Measured
3 Feb 2026
Not shown
Not other format limits such as character counts or bullet counts.
Reported by UGI Leaderboard
0.37
Style adherence · rank 90 of 370
Unit
score from 0 to 1, higher is better
Configuration
glm-4.7-flash
Measured
3 Feb 2026
Not shown
Not brand voice on your own examples: UGI's prompts are private and lean towards creative writing.
Reported by UGI Leaderboard
0.35
Style adherence · rank 166 of 370
Unit
score from 0 to 1, higher is better
Configuration
glm-4.7-flash
Measured
3 Feb 2026
Not shown
Not brand voice on your own examples: UGI's prompts are private and lean towards creative writing.
Reported by UGI Leaderboard
14.7
Writing score · rank 354 of 370
Unit
score out of 100, higher is better
Configuration
glm-4.7-flash
Measured
3 Feb 2026
Not shown
Not business copy quality: UGI's prompts are private and lean towards creative writing, and models that often refuse get no score.
Reported by UGI Leaderboard
25.0
Writing score · rank 304 of 370
Unit
score out of 100, higher is better
Configuration
glm-4.7-flash
Measured
3 Feb 2026
Not shown
Not business copy quality: UGI's prompts are private and lean towards creative writing, and models that often refuse get no score.
91.6%
Answer rate · rank 102 of 108
Unit
% of documents, higher is better
Configuration
GLM 4.7 Flash
Measured
22 Sep 2026
Not shown
Not a quality score: a low rate usually means content filters were triggered, and hallucination rates are measured on answered documents only.
9.3%
Hallucination rate · rank 47 of 108
Unit
% of summaries, lower is better
Configuration
GLM 4.7 Flash
Measured
22 Sep 2026
Not shown
Not errors in open questions or other tasks: only summarisation, judged by Vectara's own model (HHEM-2.3), not by people, on news-style documents rather than your data.

Compare GLM 4.7 Flash with