NEW: 50+ real-time integrations - track Claude, OpenAI, AWS, OpenRouter & more. Try it free →

CostGoat Logo

CostGoat

LAST UPDATED: AUGUST 13, 2026

LLM API Pricing Comparison

Compare pricing across 371+ LLM APIs from OpenAI, Anthropic, Google, DeepSeek, Mistral, xAI, and more. Sorted by quality, price, or value score.

ComparisonValue RankingsPricing GuideSave MoneyFAQ

Pricing TLDR

  • Live pricing from $0.01/M input up to $600/M output
  • Quality scores from 0-100, based on independent benchmarks (Theozard)
  • Value score = quality per dollar of output cost, so you can spot the best value fast

Official pricing:

OpenRouter API (live pricing)

Quality Scores: Theozard

Monthly LLM API Cost Comparison

Calculate by

Input Tokens

Output Tokens

API Calls / Month

Quick Examples:

Sort:

(anthropic/claude-opus-5:batch)

Context

1.0M

Quality

100

Per 1M Tokens

In: $2.50

Out: $12.50

Value

8.0

Monthly Cost

$8.75

(anthropic/claude-opus-5)

Context

1.0M

Quality

100

Per 1M Tokens

In: $5.00

Out: $25.00

Value

4.0

Monthly Cost

$17.50

(anthropic/claude-opus-5-fast)

Context

1.0M

Quality

100

Per 1M Tokens

In: $10.00

Out: $50.00

Value

2.0

Monthly Cost

$35.00

(anthropic/claude-fable-5:batch)

Context

1.0M

Quality

98

Per 1M Tokens

In: $5.00

Out: $25.00

Value

3.9

Monthly Cost

$17.50

(anthropic/claude-fable-5)

Context

1.0M

Quality

98

Per 1M Tokens

In: $10.00

Out: $50.00

Value

2.0

Monthly Cost

$35.00

(x-ai/grok-4.6)

Context

500K

Quality

97

Per 1M Tokens

In: $2.00

Out: $6.00

Value

16.2

Monthly Cost

$5.00

(openai/gpt-5.6-sol:batch)

Context

1.1M

Quality

97

Per 1M Tokens

In: $2.50

Out: $15.00

Value

6.5

Monthly Cost

$10.00

(openai/gpt-5.6-sol)

Context

1.1M

Quality

97

Per 1M Tokens

In: $5.00

Out: $30.00

Value

3.2

Monthly Cost

$20.00

(moonshotai/kimi-k3)

Context

1.0M

Quality

95

Per 1M Tokens

In: $3.00

Out: $15.00

Value

6.3

Monthly Cost

$10.50

(qwen/qwen3.8-max)

Context

1.0M

Quality

92

Per 1M Tokens

In: $2.00

Out: $6.00

Value

15.3

Monthly Cost

$5.00

(anthropic/claude-opus-4.8:batch)

Context

1.0M

Quality

91

Per 1M Tokens

In: $2.50

Out: $12.50

Value

7.3

Monthly Cost

$8.75

(anthropic/claude-opus-4.8)

Context

1.0M

Quality

91

Per 1M Tokens

In: $5.00

Out: $25.00

Value

3.6

Monthly Cost

$17.50

(anthropic/claude-opus-4.8-fast)

Context

1.0M

Quality

91

Per 1M Tokens

In: $10.00

Out: $50.00

Value

1.8

Monthly Cost

$35.00

(meta/muse-spark-1.2)

Context

1.0M

Quality

90

Per 1M Tokens

In: $1.25

Out: $4.25

Value

21.2

Monthly Cost

$3.38

(openai/gpt-5.6-terra)

Context

1.1M

Quality

90

Per 1M Tokens

In: $1.00

Out: $6.00

Value

15.0

Monthly Cost

$4.00

(openai/gpt-5.6-terra:batch)

Context

1.1M

Quality

90

Per 1M Tokens

In: $1.00

Out: $6.00

Value

15.0

Monthly Cost

$4.00

(openai/gpt-5.5:batch)

Context

1.1M

Quality

89

Per 1M Tokens

In: $2.50

Out: $15.00

Value

5.9

Monthly Cost

$10.00

(openai/gpt-5.5)

Context

1.1M

Quality

89

Per 1M Tokens

In: $5.00

Out: $30.00

Value

3.0

Monthly Cost

$20.00

(anthropic/claude-sonnet-5:batch)

Context

1.0M

Quality

88

Per 1M Tokens

In: $1.00

Out: $5.00

Value

17.6

Monthly Cost

$3.50

(x-ai/grok-4.5)

Context

500K

Quality

88

Per 1M Tokens

In: $2.00

Out: $6.00

Value

14.7

Monthly Cost

$5.00

(anthropic/claude-sonnet-5)

Context

1.0M

Quality

88

Per 1M Tokens

In: $2.00

Out: $10.00

Value

8.8

Monthly Cost

$7.00

(anthropic/claude-opus-4.7:batch)

Context

1.0M

Quality

87

Per 1M Tokens

In: $2.50

Out: $12.50

Value

7.0

Monthly Cost

$8.75

(anthropic/claude-opus-4.7)

Context

1.0M

Quality

87

Per 1M Tokens

In: $5.00

Out: $25.00

Value

3.5

Monthly Cost

$17.50

(anthropic/claude-opus-4.7-fast)

Context

1.0M

Quality

87

Per 1M Tokens

In: $30.00

Out: $150.00

Value

0.6

Monthly Cost

$105.00

(deepseek/deepseek-v4-pro-0813)

Context

1.0M

Quality

84

Per 1M Tokens

In: $0.44

Out: $0.87

Value

96.5

Monthly Cost

$0.87

(meta/muse-spark-1.1)

Context

1.0M

Quality

84

Per 1M Tokens

In: $1.25

Out: $4.25

Value

19.8

Monthly Cost

$3.38

(openai/gpt-5.4:batch)

Context

1.1M

Quality

84

Per 1M Tokens

In: $1.25

Out: $7.50

Value

11.2

Monthly Cost

$5.00

(openai/gpt-5.4)

Context

1.1M

Quality

84

Per 1M Tokens

In: $2.50

Out: $15.00

Value

5.6

Monthly Cost

$10.00

(openai/gpt-5.6-luna)

Context

1.1M

Quality

83

Per 1M Tokens

In: $0.10

Out: $0.60

Value

138.3

Monthly Cost

$0.40

(openai/gpt-5.6-luna:batch)

Context

1.1M

Quality

83

Per 1M Tokens

In: $0.10

Out: $0.60

Value

138.3

Monthly Cost

$0.40

Spending across 315+ OpenRouter models?

Monitor your OpenRouter credits and usage in real-time.

7-day free trial, no credit card required

CostGoat desktop app showing AI agent quotas, usage costs, credit balances, and subscriptions

Best Value LLM APIs by Quality Per Dollar

Value score = quality points per $1 of output cost (per 1M tokens). Higher is better. Only models above a usable quality bar are ranked, so cheap low-quality models don't top the list.

#

1

Model

Ling-3.0-flash

Provider

inclusionai

Quality

60

Output / 1M

$0.06

Value Score

952.4

#

2

Model

Upstage: Solar Pro 4

Provider

upstage

Quality

66

Output / 1M

$0.12

Value Score

550.0

#

3

Model

DeepSeek: DeepSeek V4 Flash 0731

Provider

DeepSeek

Quality

82

Output / 1M

$0.18

Value Score

455.6

#

4

Model

DeepSeek V4 Flash Latest

Provider

~deepseek

Quality

67

Output / 1M

$0.25

Value Score

265.9

#

5

Model

DeepSeek: DeepSeek V4 Flash 0423

Provider

DeepSeek

Quality

67

Output / 1M

$0.28

Value Score

239.3

#

6

Model

Xiaomi: MiMo-V2.5

Provider

Xiaomi

Quality

60

Output / 1M

$0.28

Value Score

214.3

#

7

Model

OpenAI: GPT-5.6 Luna

Provider

OpenAI

Quality

83

Output / 1M

$0.60

Value Score

138.3

#

8

Model

OpenAI: GPT-5.6 Luna (batch)

Provider

OpenAI

Quality

83

Output / 1M

$0.60

Value Score

138.3

#

9

Model

DeepSeek: DeepSeek V3.2

Provider

DeepSeek

Quality

52

Output / 1M

$0.40

Value Score

130.0

#

10

Model

Tencent: Hy3

Provider

tencent

Quality

67

Output / 1M

$0.53

Value Score

126.9

About LLM API Pricing

What is LLM API Pricing?

LLM APIs let you integrate large language models into your applications via HTTP requests. Every major AI provider (OpenAI, Anthropic, Google, DeepSeek, Mistral, xAI) offers API access to their models with per-token pricing. You pay separately for input tokens (your prompts) and output tokens (model responses), quoted per million tokens.

  • Input vs Output Token Pricing: Input tokens (prompts, context) are cheaper because they only need to be processed once. Output tokens (completions) cost 2-5x more because each token requires a full forward pass through the model. Optimizing prompt length has the biggest impact on cost.
  • Quality-Price Tradeoff: More expensive models generally deliver higher quality. Our quality scores (0-100) let you compare directly: the top models sit near 100 while budget models score lower but cost a fraction as much. The right choice depends on how much quality your task actually needs.
  • Context Window Costs: Larger context windows let you send more data per request but increase token costs. A 200K context model processing long documents costs proportionally more in input tokens than a short chatbot interaction. Choose context size based on your actual needs.

When to Use LLM API Pricing

Different use cases call for different models. Match your quality requirements to your budget using the value score to find the optimal model.

Ideal for

  • Chatbots and conversational AI, where mid-tier models offer the best quality-to-cost balance
  • Code generation, where code-specialized models optimize for the task
  • Bulk content processing, where budget models handle volume at minimal cost
  • Complex reasoning, where premium flagship models justify their cost on hard problems
  • Prototyping, where free-tier models let you build at no cost

Not ideal for

  • Real-time applications needing sub-100ms latency (consider edge-deployed models)
  • Tasks that don't need language understanding (use traditional algorithms instead)
  • Processing sensitive data with compliance requirements (check each provider's data policies)

LLM API Monthly Cost Estimates

Hobby / Prototyping

$0-10/mo

Free tier models

< 1K requests/day

Testing & development

Startup / MVP

$50-300/mo

Mid-tier models

5-20K requests/day

Single product

Growth

$300-2,000/mo

Mix of premium & budget models

20-100K requests/day

Multiple use cases

Enterprise

$2,000+/mo

Premium models for quality-critical tasks

100K+ requests/day

Model fallback chains

5 LLM API Cost Optimization Tips

1

Use a Model Cascade

Route easy queries to cheap models and only escalate to premium ones when needed. A classifier model can decide the routing. This typically saves 60-80% versus using premium models for everything.

2

Optimize Prompt Length

Input tokens cost money. Strip unnecessary context, use concise system prompts, and avoid sending full documents when a summary suffices. A 50% reduction in prompt length = 50% savings on input costs.

3

Cache Frequent Requests

If you make similar API calls repeatedly, cache responses. Many providers also offer prompt caching features that reduce costs for repeated system prompts. Anthropic's prompt caching can save up to 90% on cached tokens.

4

Value Score Beats Lowest Price

The cheapest model isn't always the best value. A rock-bottom price with a low quality score can deliver less than a mid-price model with a much higher score. Use the value-score column to find the sweet spot for your needs.

5

Monitor Per-Model Spending

Track costs per model and per use case with CostGoat. Identify which models consume the most budget, find opportunities to downgrade specific workflows, and catch cost spikes early before they become expensive surprises.

Start Tracking Your LLM API Spending

Monitor spending across OpenAI, Anthropic, Google, OpenRouter, and other LLM providers — all from one menubar app.

7-day free trial, no credit card required

CostGoat desktop app showing AI agent quotas, usage costs, credit balances, and subscriptions

LLM API Pricing FAQ

Common questions about LLM API costs, pricing models, and how to save money

AI Pricing

Gemini API PricingClaude API PricingGoogle Veo PricingAI Cost CalculatorsReplicate API PricingOpenRouter API PricingOpenRouter Free Models
DownloadsPricingDashboardContactIssuesAffiliatesTermsPrivacy

© 2026 CostGoat. All rights reserved.

Made by Functioncraft: Redis GUI Client · SSH GUI Client

Affiliate disclosure: Some links earn CostGoat a commission or credit when you sign up — no extra cost to you.