Skip to content

GPT-4.1 vs GPT-5

Data reconciled August 15, 2026

Short answer

GPT-4.1 and GPT-5 each win on different axes — neither is simply cheaper or better across the board.

GPT-5 is 1.6× cheaper on input.

GPT-4.1 is 1.3× cheaper on output.

GPT-4.1 holds a 3.9× larger context window.

GPT-5 returns a 3.9× longer answer per call.

Reasoning mode is supported only by GPT-5.

Side by side

ParameterGPT-4.1GPT-5
ProviderOpenAIOpenAI
Input, $ per 1M tokens$2.00$1.25
Output, $ per 1M tokens$8.00$10.00
Cache read, $ per 1M$0.50$0.13
Context window1,047,576272,000
Max output32,768128,000
Visionyesyes
Toolsyesyes
Reasoningnoyes
Statusavailableavailable

What it costs on your own workload

A price gap per million tokens says little until your volumes are plugged in: on a long fixed system prompt the winner is the model with cheap cache reads, not the one with a cheap input rate. Run GPT-4.1 and GPT-5 against your own numbers.

Open the cost calculator

What is deliberately absent

We do not compare answer quality and we do not reprint third-party benchmarks. We have no quality measurements of our own, and external tables go stale faster than prices while being compiled, almost always, by someone with a stake in the result. What is here is only what we reconcile ourselves every day: prices, limits and supported capabilities.

Other comparisons

Model catalog