Hop after hop
Your request is relayed from layer to layer before a model sees it. Every hop adds time before the first token — and one more place to break.
One gateway for the AI models you build on
TokenFusion is a single OpenAI-compatible gateway to leading AI models. One API key, one prepaid balance in US dollars, and every call streamed straight through — without the marketplace detour that costs you speed, capacity and clarity.
The aggregator detour
Many teams reach models through marketplaces and routers that sit on top of other routers. Each layer is one more wait before your first token, one more shared pool to queue in, and one more place a call can fail.
Your request is relayed from layer to layer before a model sees it. Every hop adds time before the first token — and one more place to break.
Shared keys and shared pools turn someone else's traffic spike into your timeout. The throughput you tested on Monday is not there on Tuesday.
Prepaid balances that expire, rounding you can't follow, and no record of what each call actually cost you.
How TokenFusion works
Your key is verified and the call's worst case is held against your balance. Calls are held against your available balance, and a call needs at least $1.00 available to start — nothing ever runs into debt.
We pass your request straight to the model and stream every token back the moment it's generated. Nothing waits in a queue in the middle.
When the call ends you're charged at the published selling rate for the tokens it reported, and the rest of the hold is released straight back to your balance.
Built for concurrency
In a shared pool, one customer's spike becomes everyone's error. TokenFusion gives every account its own fair share of calls in flight, so no single account can take all the capacity — and a burst waits at its own gate instead of knocking everyone else over.
Each account has its own limit on calls in flight, checked before anything is held or charged.
At a limit you get an immediate 429 or 503 with
Retry-After, so your client knows exactly when to try again.
Capacity is always kept back for sign-in, top-ups and your dashboard, even when calls are at their peak.
What you get
Point your existing OpenAI SDK at one base URL. Chat, streaming and image models sit behind the same endpoint.
Create a key for any model in a click, with ready-to-paste code. Revoke one without touching the rest.
Top up from $1. Your balance never expires, and it cannot go negative.
Every call is charged at the published selling rate for its model, shown to six decimal places, with per-model totals.
Fund each team from your wallet, cap each member by the month, and see who used what.
Every model states the region it is served from. If we have not been told, we say so — we never guess.
If a model cannot be served, the call is refused and nothing is billed. You never get a made-up answer.
A receipt for every top-up and a per-call usage history you can check line by line.
Pricing
Pick what you're doing and how often. Every figure comes from our live rate card, in US dollars, worked out exactly the way a call is billed. You are charged at the published selling rate shown for the model, on the usage reported for each call. If a call comes back with no usage, it is charged on an estimate — and marked as one on your usage page.
We buy capacity in volume. Where we pay less than list, we pass part of it on — the list price is shown crossed out beside your price below, so you can check it.
Pay $500 and $540.00 lands in your balance, so every call on every model costs you less.
Your balance never expires. Nothing you top up is wasted on a deadline.
No subscription, no seats, no minimum term. Each call is charged on the tokens it reports, to the millionth of a dollar.
If a model cannot be served, the call is refused and nothing is billed.
| Model | Input /1M | Output /1M | Cost per call | Per month |
|---|---|---|---|---|
| deepseek-v3.2 · global-ec-11
DeepSeek · Premium |
$0.0022 | $0.0029 |
$0.000003
|
$0.03saves $0.01 |
| MiniMax-M2.1 · global-ec-07
MiniMax · Premium |
$0.011 | $0.0433 |
$0.000024
|
$0.24saves $0.09 |
| deepseek-v4-flash · global-ec-12
DeepSeek · Premium |
$0.0977 | $0.2447 |
$0.000171
|
$1.71saves $0.67 |
| gemini-3.1-flash-lite · global-ec-16
Google · Premium |
$0.1837 | $1.102 |
$0.000514
|
$5.14saves $2.00 |
| MiniMax-M2.5 · global-ec-08
MiniMax · Premium |
$0.2571 | $1.0286 |
$0.000566
|
$5.66saves $2.20 |
| kimi-k2.6 · global-ec-22
Moonshot · Premium |
$0.3315 | $1.3776 |
$0.000745
|
$7.45saves $17.38 |
| glm-5.1 · ap-southeast-ec-34
Zhipu · Premium |
$0.7347 | $2.9388 |
$0.001616
|
$16.16saves $6.29 |
| kimi-k2.7-code · global-ec-23
Moonshot · Premium |
$0.7957 | $3.3061 |
$0.001788
|
$17.88saves $6.95 |
| claude-haiku-4-5 · global-ec-10
Anthropic · Premium |
$0.7347 | $3.6735 |
$0.001837
|
$18.37saves $7.14 |
| glm-5.2 · global-ec-18
Zhipu · Optimized |
$0.9771 | $3.4281 |
$0.002006
|
$20.06saves $7.80 |
| deepseek-v4-pro · global-ec-13
DeepSeek · Premium |
$1.4694 | $2.9388 |
$0.002351
|
$23.51saves $9.14 |
| Kimi K2.6 · global-fw-49
Moonshot AI · Premium · 262k context |
$1.14 | $4.80 |
$0.00258 |
$25.80 |
| qwen3.7-max · global-ec-33
Alibaba · Premium |
$1.4694 | $4.4082 |
$0.002792
|
$27.92saves $10.86 |
| Kimi K3 · global-fw-36
Moonshot AI · Premium · 1M context |
$3.60 | $18.00 |
$0.009 |
$90.00 |
| qwen3.6-max-preview · global-ec-30
Alibaba · Premium |
$11.0204 | $66.1224 |
$0.030857
|
$308.57saves $120.00 |
Priced per image, per second of video or per 1,000 characters of speech, as billed — not by the tasks above.
Each figure is what one call with exactly these token counts is charged on that model; your real charge follows the tokens each call reports. Per month is one call times the number of calls. A crossed-out figure is the same call at the model's list price, shown only where your price is lower.
Any whole-dollar amount from $1 to $10,000; bigger top-ups add a bonus to your balance. No contracts, no minimum term. Your balance stays on the account for as long as the account is open.
See the full rate cardOpen an account in a minute, top up, and call the model you need with one key.