Home Tools Blog About

LLM API Pricing Comparison

In short

Compare API prices per million tokens across GPT, Claude and Gemini, sortable and searchable. Prices as of July 2026.

  • Runs in your browser
  • Nothing uploaded
  • Free, no sign-up

Compare the API cost of the main GPT, Claude and Gemini models side by side, with a chart and a quick cost estimate. Prices are per 1M tokens as of July 2026.

Lowest cost for your tokens$0

Loading chart...

ModelProviderInput /1MOutput /1MContextEst. cost

Prices per 1M tokens, USD, as of July 2026. Many providers offer batch pricing (about 50% off) and prompt caching. Always check the provider for the latest.

🛡
100% PrivateNo server uploads, ever
InstantRuns in your browser
💧
No WatermarksClean output, always
🆓
Free ForeverNo accounts, no limits

How to Use LLM API Pricing Comparison

LLM API Pricing Comparison - free online tool by All Tools Verse
LLM API Pricing Comparison. Free online tool that runs in your browser.
  1. Scan the table. Every major GPT, Claude, and Gemini model is listed with its input and output price per million tokens and its context window.
  2. Sort to rank. Click any column heading to sort by input price, output price, context, or estimated cost.
  3. Search to filter. Type a model or provider name to narrow the list to what you care about.
  4. Estimate a job. Enter your input and output token counts and the last column shows the cost for each model.

Frequently Asked Questions

What do the prices mean?

They are the list prices in US dollars per one million tokens, shown separately for input and output, which is how the providers bill.

Which model is cheapest?

It depends on your mix of input and output. For raw price per token the small models like Gemini Flash-Lite and Claude Haiku are the lowest, while flagship models cost more.

Are these prices current?

They are accurate as of July 2026. Prices change often, so treat this as a guide and confirm on the provider site.

What does the context column mean?

It is the maximum number of tokens the model can consider at once, including your prompt and its reply.

Do providers offer discounts?

Yes. Many offer batch processing at around half price and prompt caching that cuts the cost of repeated context. Check each provider for details.

Is any data sent to a server?

No. The table and the estimate run entirely in your browser.

Keep going

Related Tools

All Developer tools →
Share

Embed this tool

Add this free tool to your website. Copy and paste the code: