Compare LLM Models — Price, Context Window and Capability
Every current model side by side: price per million tokens, context window, max output, modalities and cache rates.
This table compares 25 current models from 7 providers on the specifications that decide a choice: price per million input and output tokens, cached-input rate, context window, maximum output, supported modalities and tokeniser. It ranks them by the cost of a workload with your own input-to-output ratio, because a summarisation workload and a generation workload rank models in a completely different order.
Runs in your browser- Models compared
- 25
- Prices verified
- 2026-08-01
- Typical output premium
- 4–5× the input rate
- Privacy
- Runs entirely in your browser — nothing is uploaded
- Cost
- Free, unlimited, no sign-up
Frequently asked questions
Which model is cheapest?
Why compare context windows if they are all a million now?
Are open-weight models really cheaper?
How to use the llm model comparison
- 1Select one or more providers.
- 2Choose the sort by.
- 3Turn "Hide embedding models" on or off as needed.
- 4Set the typical input.
- 5Set the typical output.
- 6The result appears immediately — copy or download it.
Sources & specifications
Embed this tool
Put the working llm model comparison on your own site. It runs in your visitors' browsers exactly as it does here — free, no account, nothing uploaded.
Share this tool
Related tools
AI API Cost CalculatorWork out what a prompt costs per call, per day and per month, and compare the same workload across every major model.Context Window CalculatorPaste a document and see which models it fits inside, how much of each window it fills, and what is left for the reply.LLM Token CounterCount tokens with the real BPE vocabulary, see every token coloured in place, and compare the count across models.Prompt Injection SanitizerStrip invisible carriers, neutralise instruction-like markup and fence untrusted content before you paste it into a prompt.LLM Stream ParserPaste a raw server-sent-event stream and get the reconstructed message, tool calls, usage and stop reason.Prompt Template GeneratorWrite a prompt once with placeholders, paste rows of data, and get every filled prompt back with its token count and cost.
Last updated
More ai & llm tools