GPT-4.5
OpenAI · Status: Active · Released 2025-02-27 · Accessibility: API access
“Epoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks' [online resource].” · epoch.ai/benchmarks · Licensed CC BY 4.0 · data last verified 2026-09-21 · Methodology · Data Sources
Xentir Snapshot
Benchmark breakdown
| Benchmark | Score | Source |
|---|---|---|
Epoch Capabilities Index (ECI)A composite index published by Epoch AI, aggregating results across multiple benchmark categories into a single capability score. | 136.8 | Epoch AI Benchmarking Hub |
SWE-bench VerifiedReal-world GitHub issue resolution tasks, restricted to the human-verified subset of SWE-bench. | Not measured | Epoch AI Benchmarking Hub |
GPQA DiamondGraduate-level multiple-choice questions in biology, physics and chemistry, written to resist search-engine lookup. | 0.687 | Epoch AI Benchmarking Hub |
MATH Level 5The hardest difficulty tier of competition mathematics problems in the MATH benchmark. | 0.786 | Epoch AI Benchmarking Hub |
OTIS Mock AIME 2024-2025Mock American Invitational Mathematics Examination problems from the OTIS problem sets. | 0.378 | Epoch AI Benchmarking Hub |
FrontierMathOriginal, unpublished research-level mathematics problems. | Not measured | Epoch AI Benchmarking Hub |
FrontierMath Tier 4The hardest difficulty tier within FrontierMath. | Not measured | Epoch AI Benchmarking Hub |
SimpleQA VerifiedShort factual questions, each with a single verifiable answer. | Not measured | Epoch AI Benchmarking Hub |
Chess PuzzlesTactical chess puzzles requiring one correct sequence of moves. | Not measured | Epoch AI Benchmarking Hub |
Pricing
No licensed per-token pricing exists for this model yet. Not shown as an estimate — see Data Sources.
History
History begins 2026-07-26 — one data point so far. A trend needs at least two.
Compare
Ranked nearby
GPT-4.5 is #136 of 334 on the overall index. The models ranked closest to it:
- #133 DeepSeek-R1-Distill-Qwen-32B · DeepSeek
- #134 Qwen3-30B-A3B-Instruct (Jul 2025) · Alibaba
- #135 GPT-4.1 · OpenAI
- #137 Qwen3-30B-A3B · Alibaba
- #138 Qwen3-8B · Alibaba
- #139 DeepSeek-V3 (Mar 2025) · DeepSeek
Latest Xentir coverage
- OpenAI Examines Connecting AI Usage to Business Value · 2026-09-20
- OpenAI Introduces Australian Youth Safety Blueprint · 2026-09-20
- Anthropic’s Claude AI Was Used to Break Into OpenAI · 2026-09-18
Sources and methodology
Scores on this page come from Epoch AI Benchmarking Hub, licensed CC BY 4.0. See the full methodology and data sources pages for how rankings, variants and ties are computed, and the full AI Model Rankings table for how GPT-4.5 compares to every other measured model.