Explore software
COMPARE / LLM APIS

LLM APIs, with the math.

Compare token pricing, prompt caching savings, and inference costs across foundation models.

Commercial API inference for generative language and reasoning models. Shared multi-tenant API endpoints, standard rate limits, and pay-as-you-go billing without committed use discounts or provisioned throughput. Fine-tuning, batch API discounts, embeddings, and audio generation are excluded.
YOUR COMPARISON, YOUR CALL

Bring your shortlist.

2 / 4 selected

Choose up to four products for a closer look. The full cost ranking stays below.

Showing 4 of 4 supported products. Find products for this comparison

Your workload

Monthly usage. Adjust the assumptions to match your product.

tokens / month

Uncached input prompt tokens processed by the model API.

Maximum 10,000,000,000 tokens / month
tokens / month

Generated completion and reasoning tokens returned by the model API.

Maximum 2,000,000,000 tokens / month
tokens / month

Prompt tokens read from context cache (system prompts, docs).

Maximum 10,000,000,000 tokens / month
Estimates are in USD before taxes. Your usage stays in this URL.
Capabilities:
Decision guideOne shared workload.
Different trade-offs.
Monthly equivalentUSD · before taxNot availableClaude 3.5 HaikuNot availableDeepSeek-R1
Payment commitmentMonthlyMonthly
Fits this workload?Outside modeled limitsOutside modeled limits
What to know

Claude 3.5 Haiku retired on February 19, 2026. This model is unavailable on the Claude API.

1 more conditions
  • Claude 3.5 Sonnet retired on October 28, 2025. This model is unavailable on the Claude API.

Current DeepSeek pricing lists V4 models. This legacy R1 model and its historical token rates are not supported by the current pricing source.

1 more conditions
  • Current DeepSeek pricing lists V4 models. This legacy V3 model and its historical token rates are not supported by the current pricing source.
Outside the estimate

This pricing model is not supported.

This pricing model is not supported.

EvidenceOfficial pricing ↗Checked 2026-10-08Official pricing ↗Checked 2026-10-08

Look ahead, before you commit.

Vary standard input tokens. Keep all other inputs fixed.

Sampled growth costs
View exact values and assumptions

Each point follows your selected plan. Automatic choices use the lowest eligible supported plan per product. Missing values mean no supported plan fits. This varies one input, not all costs with customer count. Lines connect sampled estimates; prices between those points can change in steps or at plan limits.

Standard input tokens
1,000,000 tokens / month
1,000,001 tokens / month
4,500,000 tokens / month
8,000,000 tokens / month
10,000,000 tokens / month
11,500,000 tokens / month
15,000,000 tokens / month
18,500,000 tokens / month
22,000,000 tokens / month
25,500,000 tokens / month
29,000,000 tokens / month
32,500,000 tokens / month
36,000,000 tokens / month
39,500,000 tokens / month
43,000,000 tokens / month
46,500,000 tokens / month
50,000,000 tokens / month
53,500,000 tokens / month
60,500,000 tokens / month
64,000,000 tokens / month
67,500,000 tokens / month
71,000,000 tokens / month
74,500,000 tokens / month
78,000,000 tokens / month
81,500,000 tokens / month
85,000,000 tokens / month
88,500,000 tokens / month
92,000,000 tokens / month
95,500,000 tokens / month
99,000,000 tokens / month
102,500,000 tokens / month
106,000,000 tokens / month
109,500,000 tokens / month
113,000,000 tokens / month
120,000,000 tokens / month

Your monthly comparison

2 eligible products in this loaded set · Lowest modeled cost first

USD / month equivalent
01
OpenAI API Lowest estimate

GPT-4o mini

$3.08/ month equivalent
02

Llama 3.3 70B (Groq)

$10.43/ month equivalent
2 plans outside this workload
Anthropic Claude · Claude 3.5 Haiku

Claude 3.5 Haiku retired on February 19, 2026. This model is unavailable on the Claude API. Claude 3.5 Sonnet retired on October 28, 2025. This model is unavailable on the Claude API.

DeepSeek API · DeepSeek-R1

Current DeepSeek pricing lists V4 models. This legacy R1 model and its historical token rates are not supported by the current pricing source. Current DeepSeek pricing lists V4 models. This legacy V3 model and its historical token rates are not supported by the current pricing source.

Prices reflect dated, supported plans. Matching a category does not mean matching features. How estimates work

2 products on your board · Same workload, clearer trade-offs.