TokenBill LLM cost statements

Prices updated 2026-10-03 12:01 KST

Monthly cost statement

How much do LLM APIs actually cost per month?

Currency
USD
Billing
Per month
Models tracked
15
Input $ / 1M tokens
—
Output $ / 1M tokens
—
Statement date
2026-10-03 12:01 KST

Token prices lie. The same support reply can cost 17× more on one model than another — and cache hit rates change everything. Pick your workload, pick your models, get the real number.

Usage

Workload preset

Inputs

Cache math: cached input tokens are billed at 90% off — an illustrative assumption; real cache discounts vary by provider.

Models

Itemized charges

Select at least one model to see costs.

Statement guide

How to read your statement

  1. Set your workload

    Pick a preset, or type your own numbers. Requests / month is how many API calls you make. Input tokens / request is what you send — prompts and retrieved context. Output tokens / request is what the model writes back. Cache hit rate is the share of input tokens served from cache, billed at 90% off.

  2. Choose your models

    Toggle the model chips to build your comparison set. The dot on each chip is the provider’s color — it is the same color that model wears in every chart below, so you can track it across the page.

  3. Read the bill

    Amount due is the cheapest model for your workload. You save is what you keep by not picking the most expensive one. Each itemized row shows its vs cheapest multiple — a 17× means you pay seventeen times more for the same workload. The charts show the same numbers at different scales: relative size, input/output price geometry, and how cache changes the ranking.