Pricing comparison
| Model | In | Cached | Out |
|---|---|---|---|
| gpt-5.4-mini | $0.75 | $0.075 | $4.50 |
| gpt-5-mini | $0.25 | $0.025 | $2.00 |
| Price increase | 200% | 200% | 125% |
| Model | In | Cached | Out |
|---|---|---|---|
| gpt-5.4-nano | $0.20 | $0.020 | $1.25 |
| gpt-5-nano | $0.05 | $0.005 | $0.40 |
| Price increase | 300% | 300% | 213% |
Vision Pricing - Cost token multiplier
Documentation is wrong - experimentally, currently:
| Model | Multiplier | Max Billed Image Tokens | Verified |
|---|---|---|---|
| gpt-5.4-mini | 1.2 | 1843 | API |
| gpt-5.4-nano | 1.2 | 1843 | API |
| gpt-5-mini | 1.2 | 1843 | True |
| gpt-5-nano | 1.5 | 2304 | True |
Includes the billable tokens at "high’ per image. "detail":"low" is of no effect on these “patches” AI models, which is obfuscated in documentation.
Vision docs have the wrong multiplier for gpt-5 and gpt-5.4 mini/nano - unless cost is to be stealth increased. TBD.
API call, where I capture the image-only cost, and then de-multiply it back to see if it agrees with patches formula:
| model | vision | vision_mult | chat input | calculated | responses input | calculated |
|---|---|---|---|---|---|---|
| gpt-5.4-mini | patch | 1.2 | 527 | 433 | 526 | 432 |
| gpt-5-mini | patch | 1.2 | 527 | 433 | 526 | 432 |
| gpt-5.4-nano | patch | 1.2 | 527 | 433 | 526 | 432 |
| gpt-5-nano | patch | 1.5 | 656 | 432 | 655 | 432 |
Despite being designated for “future models”, and gpt-5.4-mini/nano being in future from the original documentation, these small models disallow the larger 2500 patches vision at “high” or the “original” resolution.
Multiplier absent from documentation
Max tokens billed per image per input
| Model | Multiplier | low/high | original |
|---|---|---|---|
| gpt-5.4 | 1.2 | 3000 | 12000 |
| gpt-5.3 | 1.2 | 1843 | n/a |