Models / gpt-5-chat

OpenAI

gpt-5-chat

7 published results from 1 source. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.

Provider
OpenAI
Sources
1
Our benchmarks
0
Price
Not yet published

Reported by others

1,422
Overall · rank 106 of 177
Unit
Arena rating, higher is better
Range
1,415 to 1,428
Sample
7736 votes
Configuration
gpt-5-chat
Measured
25 Sep 2026
Not shown
Not an error rate: claims that cannot be checked on the web are skipped, and preference still carries most of the weight.
1,440
Business, management and finance · rank 61 of 402
Unit
Arena rating, higher is better
Range
1,432 to 1,449
Sample
5574 votes
Configuration
gpt-5-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,389
Creative writing · rank 94 of 407
Unit
Arena rating, higher is better
Range
1,380 to 1,399
Sample
3967 votes
Configuration
gpt-5-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,443
Expert prompts · rank 84 of 359
Unit
Arena rating, higher is better
Range
1,428 to 1,458
Sample
1512 votes
Configuration
gpt-5-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,415
Instruction following · rank 92 of 409
Unit
Arena rating, higher is better
Range
1,408 to 1,422
Sample
8160 votes
Configuration
gpt-5-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,427
Overall · rank 99 of 409
Unit
Arena rating, higher is better
Range
1,422 to 1,431
Sample
31012 votes
Configuration
gpt-5-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,407
Writing, literature and language · rank 89 of 408
Unit
Arena rating, higher is better
Range
1,399 to 1,414
Sample
6717 votes
Configuration
gpt-5-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.