Models / gpt-4.5-preview-2025-02-27
OpenAIgpt-4.5-preview-2025-02-27
6 published results from 1 source. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.
- Provider
- OpenAI
- Sources
- 1
- Our benchmarks
- 0
- Price
- Not yet published
Reported by others
1,426
Business, management and finance · rank 71 of 402
- Unit
- Arena rating, higher is better
- Range
- 1,410 to 1,441
- Sample
- 1394 votes
- Configuration
- gpt-4.5-preview-2025-02-27
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,436
Creative writing · rank 26 of 407
- Unit
- Arena rating, higher is better
- Range
- 1,424 to 1,448
- Sample
- 2618 votes
- Configuration
- gpt-4.5-preview-2025-02-27
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,429
Expert prompts · rank 90 of 359
- Unit
- Arena rating, higher is better
- Range
- 1,405 to 1,453
- Sample
- 608 votes
- Configuration
- gpt-4.5-preview-2025-02-27
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,437
Instruction following · rank 57 of 409
- Unit
- Arena rating, higher is better
- Range
- 1,428 to 1,445
- Sample
- 5501 votes
- Configuration
- gpt-4.5-preview-2025-02-27
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,445
Overall · rank 68 of 409
- Unit
- Arena rating, higher is better
- Range
- 1,439 to 1,450
- Sample
- 14547 votes
- Configuration
- gpt-4.5-preview-2025-02-27
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,448
Writing, literature and language · rank 28 of 408
- Unit
- Arena rating, higher is better
- Range
- 1,438 to 1,457
- Sample
- 4092 votes
- Configuration
- gpt-4.5-preview-2025-02-27
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.