Every frontier model.
A tenth of the price.
Claude, GPT, Gemini, Grok, DeepSeek and every other top model behind one endpoint that speaks both OpenAI and Anthropic. Paste a base URL into Cursor, Claude Code or Codex and pay per token: prepaid, capped per key, 90% below list.
https://api.tokenlowcost.com/v1
What requests cost
Simulated requests, priced at today's rates
- Cursor
- VS Code
- Claude Code
- Codex
- Cline
- Roo Code
- Continue
- OpenAI SDK
- Anthropic SDK
Ranked. Best models first.
The catalog is ordered by capability, strongest first, so the top of the list is the best model money can buy — now at a tenth of its list price.
| # | Model | Tier | Context | Input | Output | You save |
|---|---|---|---|---|---|---|
| Loading models… | ||||||
15 labs. One account.
Anthropic, OpenAI, Google, xAI, DeepSeek, Qwen and more — one balance, one bill, one base URL. Make a key for each model you want and switch by switching keys.
Built like infrastructure.
Priced like a launch promo.
One endpoint for every protocol your tools speak, with the controls you would expect from a real API provider.
One endpoint. Every protocol.
The format is detected from the path — OpenAI Chat Completions, Responses or Anthropic Messages — so the same base URL works in every tool. The /v1 prefix is optional.
A budget and an expiry on every key.
One key per project, teammate or agent. Cap what it can spend, give it a lifetime, revoke it in one click — enforced on every request.
claude-opus-5.5gpt-6-solEvery request itemized.
Input, cached input, output and the exact charge — per request, per key, per day.
Cache hits at cache rates.
Coding tools resend the same context every turn. Whatever the provider serves from its cache is billed at the cached rate.
The model you pick is the model you get.
No silent downgrades, no cheaper substitutes. Each response names the model that actually answered.
Spend per day, split by model.
Hover any day in the dashboard to see exactly which models the money went to. Every request is listed with its tokens and cost.
Prepaid. No surprises.
Top up from $5 in crypto and spend it per token. No seats, no subscription, no monthly minimum — and no overdraft: at zero, requests stop.
What would your month cost?
Pick a model and drag the sliders. The numbers use live rates, including the cached-input rate.
Two settings. Keep your tools.
- 1
Create an account and a key
Pick a model. You can also set a spending limit and an expiry date. The key is shown only once, so copy it right away.
- 2
Set up your tool
Every tool uses the same base URL,
https://api.tokenlowcost.com/v1. Claude Code and Codex need one pasted line. In Cursor and VS Code, you paste your key in the settings. - 3
Keep working
Your tools work as before. The dashboard shows every request and what it cost.
Step-by-step guides for every tool are in the documentation.
Straight answers.
Setup for every tool is in the docs. Current rates for every model are on the pricing page.
How can it be this cheap?
Is there free credit to start?
Do I really get the model I choose?
Why is each key tied to one model?
Are cached tokens cheaper?
How does billing work?
Do you store my prompts?
Which tools are supported?
How do I add funds?
Same models.
A tenth of the bill.
Create an account and get $50 of credit to start. Make a key, paste the base URL: setup takes about a minute.