How it works
Built so anyone can choose well.
BetterLLMs is a decision helper, not another chat app. It answers three questions: which model, how much it costs, and whether switching is worth it.
01
Name the task
Coding, writing, analysis, long documents, creative work, cheap bulk jobs, or general chat.
02
See three picks
Best overall, best value, and cheapest usable. You choose the tradeoff — we don’t hide it.
03
Check the bill
Cost is explained in everyday language: what you send vs what the model writes, plus an example chat in dollars. Open Advanced if you want the per-million-token rates.
04
Switch with care
Changing models mid-thread can drop context. We say Switch or Don’t switch in one line.
05
Plan Copilot credits
Enter your monthly credits and how many you’ve used (leave used empty for 0). We pace what’s left by task: everyday model, harder model, and a cheaper one.
06
Read the credit tips
Separate tabs for Auto, Chat, agents, pasted context, and which model to open — so leftover credits last the month.
What “$2/M in · $10/M out” means
That shorthand is industry jargon. “In” is reading your message. “Out” is writing the answer. “/M” means per million tokens — about a 1,500-page book of text, not one chat. A typical coding session on Claude Sonnet 5 is a few cents, not $2 or $10.
Catalog sources
19 current models across 8 providers, plus earlier versions for comparison. This is a curated text and coding catalog, not every model or specialized media API.
Latest additions checked on . Prices are USD per million tokens at base text rates, not live quotes. Previous entries retain historical snapshots. Current means a current recommendation, not guaranteed availability in every app or region. Qwen uses international list rates. Cache writes, reasoning, tools, and long-context tiers can change the total bill. Speed and task-fit descriptions are editorial guidance, not measured benchmarks.
- Anthropic models and pricing
- OpenAI models and pricing
- Google models
- Google pricing
- DeepSeek models and pricing
- xAI models and pricing
- Kimi models and pricing
- Mistral Medium 3.5
- Mistral Small 4
- Qwen models
- Qwen international pricing
Task picks are editorial suggestions, not benchmark rankings. The Copilot planner uses a separate model list; API availability does not confirm Copilot availability. Confirm your plan and organization settings.
What we don’t do yet
No accounts, no live price sync, no auto-routing, and no editor plugins. Those come later. This MVP is here to make the choice obvious.