NEW: 70+ real-time integrations - track Claude, OpenAI, AWS, OpenRouter & more. Try it free →

CostGoat Logo

CostGoat

LAST UPDATED: SEPTEMBER 16, 2026

LLM API Pricing Comparison

Compare pricing across 400+ LLM APIs from OpenAI, Anthropic, Google, DeepSeek, Mistral, xAI, and more. Sorted by quality, price, or value score.

ComparisonValue RankingsPricing GuideSave MoneyFAQ

Pricing TLDR

  • Live pricing from $0.02/M input up to $600/M output
  • Quality scores from 0-100, based on independent benchmarks (Theozard)
  • Value score = quality per dollar of output cost, so you can spot the best value fast

Official pricing:

OpenRouter API (live pricing)

Quality Scores: Theozard

Monthly LLM API Cost Comparison

Calculate by

Input Tokens

Output Tokens

API Calls / Month

Quick Examples:

Sort:

(anthropic/claude-fable-5.1:batch)

Context

1.0M

Quality

100

Per 1M Tokens

In: $5.00

Out: $25.00

Value

4.0

Monthly Cost

$17.50

(anthropic/claude-fable-5.1)

Context

1.0M

Quality

100

Per 1M Tokens

In: $10.00

Out: $50.00

Value

2.0

Monthly Cost

$35.00

(openai/gpt-6-astra:batch)

Context

1.1M

Quality

99

Per 1M Tokens

In: $5.00

Out: $25.00

Value

4.0

Monthly Cost

$17.50

(openai/gpt-6-astra)

Context

1.1M

Quality

99

Per 1M Tokens

In: $10.00

Out: $50.00

Value

2.0

Monthly Cost

$35.00

(anthropic/claude-opus-5:batch)

Context

1.0M

Quality

95

Per 1M Tokens

In: $2.50

Out: $12.50

Value

7.6

Monthly Cost

$8.75

(anthropic/claude-opus-5)

Context

1.0M

Quality

95

Per 1M Tokens

In: $5.00

Out: $25.00

Value

3.8

Monthly Cost

$17.50

(anthropic/claude-fable-5:batch)

Context

1.0M

Quality

93

Per 1M Tokens

In: $5.00

Out: $25.00

Value

3.7

Monthly Cost

$17.50

(anthropic/claude-fable-5)

Context

1.0M

Quality

93

Per 1M Tokens

In: $10.00

Out: $50.00

Value

1.9

Monthly Cost

$35.00

(meta/muse-spark-1.3)

Context

1.0M

Quality

90

Per 1M Tokens

In: $1.25

Out: $4.25

Value

21.2

Monthly Cost

$3.38

(openai/gpt-5.6-sol:batch)

Context

1.1M

Quality

88

Per 1M Tokens

In: $1.00

Out: $5.00

Value

17.6

Monthly Cost

$3.50

(openai/gpt-5.6-sol)

Context

1.1M

Quality

88

Per 1M Tokens

In: $2.00

Out: $10.00

Value

8.8

Monthly Cost

$7.00

(qwen/qwen3.8-max-0902)

Context

1.0M

Quality

85

Per 1M Tokens

In: $2.00

Out: $6.00

Value

14.2

Monthly Cost

$5.00

(z-ai/glm-5.3:batch)

Context

1.0M

Quality

84

Per 1M Tokens

In: $0.70

Out: $2.20

Value

38.2

Monthly Cost

$1.80

(z-ai/glm-5.3)

Context

1.3M

Quality

84

Per 1M Tokens

In: $1.40

Out: $4.40

Value

19.1

Monthly Cost

$3.60

(x-ai/grok-4.6)

Context

500K

Quality

83

Per 1M Tokens

In: $2.00

Out: $6.00

Value

13.8

Monthly Cost

$5.00

(moonshotai/kimi-k3)

Context

1.0M

Quality

82

Per 1M Tokens

In: $2.65

Out: $13.28

Value

6.2

Monthly Cost

$9.29

(moonshotai/kimi-k3:batch)

Context

1.0M

Quality

82

Per 1M Tokens

In: $3.00

Out: $15.00

Value

5.5

Monthly Cost

$10.50

(z-ai/glm-5.3-flash:batch)

Context

1.0M

Quality

79

Per 1M Tokens

In: $0.08

Out: $0.25

Value

316.0

Monthly Cost

$0.20

(z-ai/glm-5.3-flash)

Context

1.3M

Quality

79

Per 1M Tokens

In: $0.10

Out: $0.33

Value

237.0

Monthly Cost

$0.27

(openai/gpt-5.6-terra:batch)

Context

1.1M

Quality

79

Per 1M Tokens

In: $1.00

Out: $6.00

Value

13.2

Monthly Cost

$4.00

(openai/gpt-5.6-terra)

Context

1.1M

Quality

79

Per 1M Tokens

In: $2.00

Out: $12.00

Value

6.6

Monthly Cost

$8.00

(anthropic/claude-opus-4.8:batch)

Context

1.0M

Quality

79

Per 1M Tokens

In: $2.50

Out: $12.50

Value

6.3

Monthly Cost

$8.75

(anthropic/claude-opus-4.8)

Context

1.0M

Quality

79

Per 1M Tokens

In: $5.00

Out: $25.00

Value

3.2

Monthly Cost

$17.50

(google/gemini-3.8-flash:batch)

Context

1.0M

Quality

77

Per 1M Tokens

In: $0.38

Out: $1.88

Value

41.1

Monthly Cost

$1.31

(google/gemini-3.8-flash)

Context

1.0M

Quality

77

Per 1M Tokens

In: $0.75

Out: $3.75

Value

20.5

Monthly Cost

$2.62

(anthropic/claude-opus-4.7:batch)

Context

1.0M

Quality

76

Per 1M Tokens

In: $2.50

Out: $12.50

Value

6.1

Monthly Cost

$8.75

(anthropic/claude-opus-4.7)

Context

1.0M

Quality

76

Per 1M Tokens

In: $5.00

Out: $25.00

Value

3.0

Monthly Cost

$17.50

(meta/muse-spark-1.2)

Context

1.0M

Quality

75

Per 1M Tokens

In: $1.25

Out: $4.25

Value

17.6

Monthly Cost

$3.38

(qwen/qwen3.8-2.4t-a95b)

Context

1.0M

Quality

75

Per 1M Tokens

In: $2.00

Out: $6.00

Value

12.5

Monthly Cost

$5.00

(qwen/qwen3.8-2.4t-a95b:batch)

Context

1.0M

Quality

75

Per 1M Tokens

In: $2.00

Out: $6.00

Value

12.5

Monthly Cost

$5.00

Spending across 315+ OpenRouter models?

Monitor your OpenRouter credits and usage in real-time.

Free 7-day trial. No sign-up, no credit card.

CostGoat desktop app showing AI agent quotas, usage costs, credit balances, and subscriptions

Best Value LLM APIs by Quality Per Dollar

Value score = quality points per $1 of output cost (per 1M tokens). Higher is better. Only models above a usable quality bar are ranked, so cheap low-quality models don't top the list.

#

1

Model

DeepSeek: DeepSeek V4 Flash 0731

Provider

DeepSeek

Quality

65

Output / 1M

$0.12

Value Score

541.7

#

2

Model

Z.ai: GLM 5.3 Flash (batch)

Provider

Z.ai

Quality

79

Output / 1M

$0.25

Value Score

316.0

#

3

Model

Z.ai: GLM 5.3 Flash

Provider

Z.ai

Quality

79

Output / 1M

$0.33

Value Score

237.0

#

4

Model

DeepSeek: DeepSeek V4 Flash Vision Exp (batch)

Provider

DeepSeek

Quality

66

Output / 1M

$0.33

Value Score

200.0

#

5

Model

DeepSeek: DeepSeek V4 Flash 0731 (batch)

Provider

DeepSeek

Quality

65

Output / 1M

$0.33

Value Score

197.0

#

6

Model

Upstage: Solar Pro 4

Provider

upstage

Quality

53

Output / 1M

$0.36

Value Score

147.2

#

7

Model

DeepSeek: DeepSeek V4.1 Flash

Provider

DeepSeek

Quality

74

Output / 1M

$0.60

Value Score

123.3

#

8

Model

OpenAI: GPT-5.6 Luna (batch)

Provider

OpenAI

Quality

70

Output / 1M

$0.60

Value Score

116.7

#

9

Model

DeepSeek: DeepSeek V4 Flash Vision Exp

Provider

DeepSeek

Quality

66

Output / 1M

$0.66

Value Score

100.0

#

10

Model

OpenAI: GPT-5.6 Luna

Provider

OpenAI

Quality

70

Output / 1M

$1.20

Value Score

58.3

About LLM API Pricing

What is LLM API Pricing?

LLM APIs let you integrate large language models into your applications via HTTP requests. Every major AI provider (OpenAI, Anthropic, Google, DeepSeek, Mistral, xAI) offers API access to their models with per-token pricing. You pay separately for input tokens (your prompts) and output tokens (model responses), quoted per million tokens.

  • Input vs Output Token Pricing: Input tokens (prompts, context) are cheaper because they only need to be processed once. Output tokens (completions) cost 2-5x more because each token requires a full forward pass through the model. Optimizing prompt length has the biggest impact on cost.
  • Quality-Price Tradeoff: More expensive models generally deliver higher quality. Our quality scores (0-100) let you compare directly: the top models sit near 100 while budget models score lower but cost a fraction as much. The right choice depends on how much quality your task actually needs.
  • Context Window Costs: Larger context windows let you send more data per request but increase token costs. A 200K context model processing long documents costs proportionally more in input tokens than a short chatbot interaction. Choose context size based on your actual needs.

When to Use LLM API Pricing

Different use cases call for different models. Match your quality requirements to your budget using the value score to find the optimal model.

Ideal for

  • Chatbots and conversational AI, where mid-tier models offer the best quality-to-cost balance
  • Code generation, where code-specialized models optimize for the task
  • Bulk content processing, where budget models handle volume at minimal cost
  • Complex reasoning, where premium flagship models justify their cost on hard problems
  • Prototyping, where free-tier models let you build at no cost

Not ideal for

  • Real-time applications needing sub-100ms latency (consider edge-deployed models)
  • Tasks that don't need language understanding (use traditional algorithms instead)
  • Processing sensitive data with compliance requirements (check each provider's data policies)

LLM API Monthly Cost Estimates

Hobby / Prototyping

$0-10/mo

Free tier models

< 1K requests/day

Testing & development

Startup / MVP

$50-300/mo

Mid-tier models

5-20K requests/day

Single product

Growth

$300-2,000/mo

Premium & budget mix

20-100K requests/day

Multiple use cases

Enterprise

$2,000+/mo

Premium models

100K+ requests/day

Model fallback chains

5 LLM API Cost Optimization Tips

1

Use a Model Cascade

Route easy queries to cheap models and only escalate to premium ones when needed. A classifier model can decide the routing. This typically saves 60-80% versus using premium models for everything.

2

Optimize Prompt Length

Input tokens cost money. Strip unnecessary context, use concise system prompts, and avoid sending full documents when a summary suffices. A 50% reduction in prompt length = 50% savings on input costs.

3

Cache Frequent Requests

If you make similar API calls repeatedly, cache responses. Many providers also offer prompt caching features that reduce costs for repeated system prompts. Anthropic's prompt caching can save up to 90% on cached tokens.

4

Value Score Beats Lowest Price

The cheapest model isn't always the best value. A rock-bottom price with a low quality score can deliver less than a mid-price model with a much higher score. Use the value-score column to find the sweet spot for your needs.

5

Monitor Per-Model Spending

Track costs per model and per use case with CostGoat. Identify which models consume the most budget, find opportunities to downgrade specific workflows, and catch cost spikes early before they become expensive surprises.

Start Tracking Your LLM API Spending

Monitor spending across OpenAI, Anthropic, Google, OpenRouter, and other LLM providers, all from one menubar app.

Free 7-day trial. No sign-up, no credit card.

CostGoat desktop app showing AI agent quotas, usage costs, credit balances, and subscriptions

LLM API Pricing FAQ

Common questions about LLM API costs, pricing models, and how to save money

AI Pricing

Gemini API PricingClaude API PricingGoogle Veo PricingAI Cost CalculatorsReplicate API PricingOpenRouter API PricingOpenRouter Free Models
DownloadsPricingDealsAccountContactIssuesAffiliatesTermsPrivacy

© 2026 CostGoat. All rights reserved.

Made by Functioncraft: Redis GUI Client · SSH GUI Client · Code Signing Costs

Affiliate disclosure: Some links earn CostGoat a commission or credit when you sign up, at no extra cost to you.