OpenAI

GPT-4.1 Pricing

GPT-4.1 is a premium-tier language model from OpenAI, priced at $2.00 per 1M input tokens and $8.00 per 1M output tokens.

Input
$2.00
per 1M tokens
Output
$8.00
per 1M tokens
Cached input
$0.50
per 1M tokens 75% off
Context
1000K
tokens
Max output
32K
tokens

What does GPT-4.1 cost in practice?

These are real-world monthly estimates based on common workload sizes. They assume a 2:1 input-to-output ratio (typical of chat + retrieval-augmented apps).

Workload Input / mo Output / mo Monthly cost Annualized
Light prototyping 0.1M tokens 0.1M tokens $0.60 $7.20
Medium production 5.0M tokens 2.5M tokens $30.00 $360.00
Heavy production 50.0M tokens 25.0M tokens $300.00 $3600.00

Need to model your own usage? Use the LLM pricing calculator →

GPT-4.1 features

Alternatives to GPT-4.1

Cheaper, same-tier alternatives

If GPT-4.1 is more than you need, these premium-tier models cost less per token:

Higher-priced same-tier models

If you've outgrown GPT-4.1, consider these:

Need the cheapest OpenAI option? GPT-5 Nano starts at $ 0.05 input / $ 0.40 output per 1M tokens →

FAQ

How much does GPT-4.1 cost per 1M tokens?

GPT-4.1 costs $2.00 per 1M input tokens and $8.00 per 1M output tokens. Cached input drops to $0.50 per 1M tokens — a 75% discount on repeated context.

What is GPT-4.1's context window?

GPT-4.1 supports a 1,000,000-token context window with a maximum output of 32,000 tokens per response.

What features does GPT-4.1 support?

Vision, Tool Use, Streaming.

How much does GPT-4.1 cost for medium production usage?

For ~5M input + 2.5M output tokens per month, GPT-4.1 costs approximately $30.00/mo (≈$ 360.00/yr).

What are cheaper alternatives to GPT-4.1?

Same-tier alternatives that cost less: Mistral Large 3 .