Quick answer: GPT-4.5 (API: gpt-4-5-preview) is OpenAI's February 2025 large-scale GPT-4 model, scoring 88% on HumanEval and 5.44% on Humanity's Last Exam. It was OpenAI's largest dense model at launch, trained with a focus on world knowledge and emotional intelligence. Priced at $75 input / $150 output per 1M tokens — extremely expensive relative to performance; most workloads are better served by GPT-4.1 or GPT-5.x.
Where GPT-4.5 leads
Where it lags
Best for: Research contexts requiring GPT-4.5-specific outputs; teams that require the specific model version for reproducibility.
GPT-4.5 (gpt-4-5-preview) was OpenAI's last major GPT-4-series model before the GPT-5 launch. Released February 27, 2025, it was OpenAI's largest dense language model — trained at greater scale than GPT-4o with an emphasis on world knowledge depth and more natural, human-like conversation rather than raw reasoning scores.
OpenAI positioned GPT-4.5 as a research preview of large-scale dense training, in contrast to the o-series reasoning models and later GPT-5. Its pricing ($75/$150 per 1M) reflected its experimental and large-scale nature rather than production economics — at that price, almost every workload is more cost-effectively served by GPT-4o, GPT-4.1, or a GPT-5 family model.
For teams not explicitly requiring the gpt-4-5-preview model ID, there is no practical reason to use GPT-4.5 over GPT-5 mini or GPT-4.1, both of which offer better benchmark performance at a fraction of the cost.
| Field | Value |
|---|---|
| Organization | OpenAI |
| Parameters | Undisclosed (largest dense GPT-4 model) |
| Context window | 128,000 tokens |
| Max output | 16,384 tokens |
| API model ID | gpt-4-5-preview |
| License | Proprietary (API only) |
| Release date | February 27, 2025 |
| Knowledge cutoff | June 2024 (estimated) |
| Modality | Text + Vision (multimodal) |
| Input (per 1M tokens) | Output (per 1M tokens) | |
|---|---|---|
| OpenAI API | $75.00 | $150.00 |
Pricing per OpenAI pricing page. This is one of the highest per-token rates OpenAI has offered — verify current pricing and consider GPT-4.1 ($2/$8) or GPT-5 mini as alternatives.
GPT-4.5 has a 128,000-token context window — roughly 90 pages of text in a single request.
| Benchmark | Score | Source | Date |
|---|---|---|---|
| HumanEval | 88.0% | Benchgen evaluation | 2025-07 |
| Humanity's Last Exam | 5.44% | Benchgen evaluation | 2025-07 |
| SHADE-Arena | 0.0 overall success | Anthropic research post | 2025-06 |
| Model | Context | HumanEval | HLE | Price (in/out per 1M) |
|---|---|---|---|---|
| GPT-4.5 | 128K | 88.0% | 5.44% | $75 / $150 |
| GPT-4.1 | 1M | — | 5.40% | $2 / $8 |
| o3 | 200K | — | 20.32% | $10 / $40 |
| Gemini 2.5 Pro | 1M | — | 21.64% | $1.25 / $10 |
GPT-4.5's HumanEval score (88%) is competitive, but GPT-4.1 achieves a similar HLE score at $2/$8 — a 97% cost saving. For almost all production use cases, migrating away from GPT-4.5 is strongly recommended.
Specs and scores sourced from OpenAI's official GPT-4.5 announcement (February 2025) and Benchgen evaluations. Pricing cited to the OpenAI pricing page. Last updated 2026-07-24.
This model isn’t on any benchmark leaderboard yet.