How to estimate LLM API cost per month
Multiply three numbers: calls a day, tokens per call, and the model’s price per million tokens, once for input and once for output. A chat model bills both. A classifier or a moderation check sends a few hundred tokens in and gets a few back, so input dominates. A chatbot reply is the reverse: output costs several times more per token and there is more of it.
The prices this LLM cost calculator uses
US dollars per million tokens, standard tier, copied from each vendor on 2026-10-02. Click a model for its source.
| Model | Input | Output |
|---|---|---|
| Jev TypeSafe | $0.042 | $0 |
| GPT-6 Luna OpenAI | $0.1 | $0.5 |
| GPT-4o mini OpenAI | $0.15 | $0.6 |
| GPT-5.4 mini OpenAI | $0.75 | $4.5 |
| GPT-6.1 Sol OpenAI | $2 | $10 |
| Gemini 3.5 Flash-Lite Google | $0.3 | $2.5 |
| Gemini 3.8 Flash Google | $0.75 | $3.75 |
| Claude Haiku 4.5 Anthropic | $1 | $5 |
| Claude Sonnet 5.5 Anthropic | $2 | $10 |
| Claude Opus 5.5 Anthropic | $4 | $20 |
LLM cost calculator questions
- How is the cost worked out?
- Cost per call is input tokens times the model's input price, plus output tokens times its output price, divided by a million. The month is calls a day times 30. It uses each vendor's list price, with no discounts, no prompt caching and no batch pricing.
- Why is Jev so much cheaper?
- Jev bills input tokens only, at $0.042 per million, and charges nothing for output, because it returns a typed answer (yes or no, a choice, a score) instead of text. A chat model bills for every token it writes. When your task is a decision, that difference is most of the bill.
- Can Jev replace the other models here?
- Only for decisions. It cannot write, summarize or answer in prose. Tick the box that says your calls need written text, and Jev leaves the comparison. For the jobs it does suit, see how it compares in the published runs on the Jev vs an LLM page.
- Are the prices current?
- They were copied from each vendor's pricing page on 2026-10-02, and each model name links to its source. Vendors change prices, and Google lists a rise for Gemini 3.8 Flash from 1 January 2027, so check the source before you commit a budget.
- Why do the same tokens cost differently by model?
- Tokenizers differ. The same text can be a different number of tokens on different models, and Anthropic says its newer tokenizer produces about 30% more tokens for the same text. The calculator takes your token counts at face value, so measure them on each model you compare.
- Does this tool call a model or save anything?
- No. It is arithmetic in your browser. The numbers live only in the page link, so you can share them.
What Jev costs on real work
The calculator uses list prices. For what people actually paid, with the volume and the bill they published, see Jev pricing and Jev vs an LLM.