Models / gpt-5.2-chat-latest-20260210

OpenAI

gpt-5.2-chat-latest-20260210

7 published results from 1 source. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.

Provider
OpenAI
Sources
1
Our benchmarks
0
Price
Not yet published

Reported by others

1,451
Overall · rank 47 of 177
Unit
Arena rating, higher is better
Range
1,447 to 1,455
Sample
35639 votes
Configuration
gpt-5.2-chat-latest-20260210
Measured
25 Sep 2026
Not shown
Not an error rate: claims that cannot be checked on the web are skipped, and preference still carries most of the weight.
1,485
Business, management and finance · rank 6 of 402
Unit
Arena rating, higher is better
Range
1,477 to 1,493
Sample
7020 votes
Configuration
gpt-5.2-chat-latest-20260210
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,433
Creative writing · rank 36 of 407
Unit
Arena rating, higher is better
Range
1,425 to 1,442
Sample
5636 votes
Configuration
gpt-5.2-chat-latest-20260210
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,491
Expert prompts · rank 21 of 359
Unit
Arena rating, higher is better
Range
1,480 to 1,502
Sample
3158 votes
Configuration
gpt-5.2-chat-latest-20260210
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,457
Instruction following · rank 33 of 409
Unit
Arena rating, higher is better
Range
1,450 to 1,463
Sample
11744 votes
Configuration
gpt-5.2-chat-latest-20260210
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,476
Overall · rank 19 of 409
Unit
Arena rating, higher is better
Range
1,472 to 1,480
Sample
35962 votes
Configuration
gpt-5.2-chat-latest-20260210
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,453
Writing, literature and language · rank 26 of 408
Unit
Arena rating, higher is better
Range
1,446 to 1,460
Sample
8487 votes
Configuration
gpt-5.2-chat-latest-20260210
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.