LLM Pricing Calculator

Compare pricing across 25+ LLM models from Anthropic, OpenAI, Google, Meta, Mistral, DeepSeek, and xAI. Calculate cost per request, daily, and monthly estimates.

Usage
Providers
Cheapest / month$0.750GPT-5 Nano
Most expensive / month$315.00GPT-5.2 Pro
Models compared23

Save your results — get the cheatsheet

Drop your email and I'll send over my Prompt Playbook (free PDF). You'll also join the newsletter — unsubscribe anytime.

Or preview the Prompt Playbook landing page →

Monthly Cost Comparison

GPT-5 Nano
$0.750
Mistral Small 3.1
$0.750
Llama 3.3 70B
$0.810
DeepSeek-V3
$0.840
Gemini 2.0 Flash
$0.900
Llama 4 Scout
$1.20
DeepSeek-V3.2
$1.47
Grok-3 Mini
$1.65
Llama 4 Maverick
$1.94
Claude Haiku 4.5
$2.63
GPT-4.1 Mini
$3.60
GPT-5 Mini
$3.75
Mistral Medium 3
$4.20
Gemini 2.5 Flash
$4.65
DeepSeek-R1
$4.94
Mistral Large 3
$15.00
GPT-4.1
$18.00
Gemini 2.5 Pro
$18.75
GPT-5.2
$26.25
Claude Sonnet 4.6
$31.50
Grok-3
$31.50
Claude Opus 4.6
$52.50
GPT-5.2 Pro
$315.00
ModelProviderInput $/1M Output $/1M Context Per RequestDailyMonthly Features
GPT-5 NanoOpenAI$0.05$0.4128K<$0.001$0.025$0.750Tool UseStreaming
Mistral Small 3.1Mistral$0.1$0.3128K<$0.001$0.025$0.750VisionTool UseStreaming
Llama 3.3 70B*Meta$0.12$0.3128K<$0.001$0.027$0.810Tool UseStreaming
DeepSeek-V3DeepSeek$0.14$0.28128K<$0.001$0.028$0.840Tool UseStreaming
Gemini 2.0 FlashGoogle$0.1$0.41.0M<$0.001$0.030$0.900VisionTool UseStreamingAudio
Llama 4 Scout*Meta$0.15$0.510.0M<$0.001$0.040$1.20VisionTool UseStreaming
DeepSeek-V3.2*DeepSeek$0.28$0.42128K<$0.001$0.049$1.47Tool UseStreamingReasoning
Grok-3 MinixAI$0.3$0.51.0M<$0.001$0.055$1.65Tool UseStreamingReasoning
Llama 4 Maverick*Meta$0.22$0.851.0M<$0.001$0.065$1.94VisionTool UseStreaming
Claude Haiku 4.5Anthropic$0.25$1.25200K<$0.001$0.088$2.63VisionTool UseStreaming
GPT-4.1 MiniOpenAI$0.4$1.61.0M$0.0012$0.120$3.60VisionTool UseStreaming
GPT-5 MiniOpenAI$0.25$2128K$0.0013$0.125$3.75VisionTool UseStreaming
Mistral Medium 3Mistral$0.4$2128K$0.0014$0.140$4.20VisionTool UseStreaming
Gemini 2.5 FlashGoogle$0.3$2.51.0M$0.0015$0.155$4.65VisionTool UseStreamingReasoningAudio
DeepSeek-R1DeepSeek$0.55$2.19128K$0.0016$0.165$4.94Tool UseStreamingReasoning
Mistral Large 3Mistral$2$6256K$0.0050$0.500$15.00VisionTool UseStreaming
GPT-4.1OpenAI$2$81.0M$0.0060$0.600$18.00VisionTool UseStreaming
Gemini 2.5 Pro*Google$1.25$101.0M$0.0063$0.625$18.75VisionTool UseStreamingReasoningAudioVideo
GPT-5.2OpenAI$1.75$14400K$0.0088$0.875$26.25VisionTool UseStreamingWeb Search
Claude Sonnet 4.6Anthropic$3$15200K$0.010$1.05$31.50VisionTool UseStreamingReasoningComputer Use
Grok-3xAI$3$151.0M$0.010$1.05$31.50VisionTool UseStreamingReasoning
Claude Opus 4.6Anthropic$5$25200K$0.018$1.75$52.50VisionTool UseStreamingReasoningComputer Use
GPT-5.2 ProOpenAI$21$168400K$0.105$10.50$315.00VisionTool UseStreamingReasoningWeb Search

Prices as of March 2026. Verify on each provider's official pricing page before making decisions. Llama model prices are representative averages across hosted providers.

What the LLM pricing calculator does

Every provider prices input and output tokens differently, and the gap between the cheapest and most capable models is often 20–50×. This calculator compares 25+ models from Anthropic, OpenAI, Google, Meta, Mistral, DeepSeek, and xAI — enter your input and output token counts (or requests per day) and it shows cost per request, per day, and per month side by side, so you can size a workload before you build it.

How to estimate your cost

  1. Set your average input and output tokens per request (a typical chat turn is a few hundred to a few thousand).
  2. Enter your request volume — per request, per day, or per month.
  3. Compare the monthly total across models to find the cheapest one that meets your quality bar.

Why input vs output pricing matters

Output tokens usually cost 3–5× more than input tokens, so a workload that reads a lot but writes a little (classification, extraction) has a very different cost profile than one that generates long responses. Modelling both separately — as this tool does — is the difference between a realistic budget and a surprise bill. For live per-model figures, the site also maintains an LLM pricing tracker.

Frequently asked questions

How many tokens is my prompt?

Roughly ¾ of a word per token in English — 1,000 tokens ≈ 750 words. For an exact count, use the token counter.

Are the prices current?

The calculator uses the site's maintained pricing data across the major providers; the pricing tracker carries the live per-model figures.

Is it free?

Yes — a free browser tool, no signup.

SharePost

More tools like this