Models / gpt-4.5-preview-2025-02-27

OpenAI

gpt-4.5-preview-2025-02-27

6 published results from 1 source. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.

Provider
OpenAI
Sources
1
Our benchmarks
0
Price
Not yet published

Reported by others

1,426
Business, management and finance · rank 71 of 402
Unit
Arena rating, higher is better
Range
1,410 to 1,441
Sample
1394 votes
Configuration
gpt-4.5-preview-2025-02-27
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,436
Creative writing · rank 26 of 407
Unit
Arena rating, higher is better
Range
1,424 to 1,448
Sample
2618 votes
Configuration
gpt-4.5-preview-2025-02-27
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,429
Expert prompts · rank 90 of 359
Unit
Arena rating, higher is better
Range
1,405 to 1,453
Sample
608 votes
Configuration
gpt-4.5-preview-2025-02-27
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,437
Instruction following · rank 57 of 409
Unit
Arena rating, higher is better
Range
1,428 to 1,445
Sample
5501 votes
Configuration
gpt-4.5-preview-2025-02-27
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,445
Overall · rank 68 of 409
Unit
Arena rating, higher is better
Range
1,439 to 1,450
Sample
14547 votes
Configuration
gpt-4.5-preview-2025-02-27
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,448
Writing, literature and language · rank 28 of 408
Unit
Arena rating, higher is better
Range
1,438 to 1,457
Sample
4092 votes
Configuration
gpt-4.5-preview-2025-02-27
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.