TokenFusion

One gateway for the AI models you build on

Every model.
One key.
No detours.

TokenFusion is a single OpenAI-compatible gateway to leading AI models. One API key, one prepaid balance in US dollars, and every call streamed straight through — without the marketplace detour that costs you speed, capacity and clarity.

  • OpenAI-compatible
  • Pay per token
  • Balance never expires
  • No contracts, no minimum term

The aggregator detour

Aggregators made AI access simple.
Then came the lag, the queues and the 429s.

Many teams reach models through marketplaces and routers that sit on top of other routers. Each layer is one more wait before your first token, one more shared pool to queue in, and one more place a call can fail.

Your app The model Marketplace Router Shared pool Reseller TokenFusion · one hop
A typical aggregated route: tokens bunch at every hop, some never arrive TokenFusion: one gateway, streamed straight through
Latency

Hop after hop

Your request is relayed from layer to layer before a model sees it. Every hop adds time before the first token — and one more place to break.

Concurrency

Everyone in one pool

Shared keys and shared pools turn someone else's traffic spike into your timeout. The throughput you tested on Monday is not there on Tuesday.

Billing

A bill you can't check

Prepaid balances that expire, rounding you can't follow, and no record of what each call actually cost you.

How TokenFusion works

One hop. Streamed straight through.
Settled on what the call really used.

Your app TokenFusion gateway The model Your balance
  1. 1Checked before it runs

    Your key is verified and the call's worst case is held against your balance. Calls are held against your available balance, and a call needs at least $1.00 available to start — nothing ever runs into debt.

  2. 2One gateway hop, streamed live

    We pass your request straight to the model and stream every token back the moment it's generated. Nothing waits in a queue in the middle.

  3. 3Settled on real usage

    When the call ends you're charged at the published selling rate for the tokens it reported, and the rest of the hold is released straight back to your balance.

Built for concurrency

Your traffic keeps its own lane.

In a shared pool, one customer's spike becomes everyone's error. TokenFusion gives every account its own fair share of calls in flight, so no single account can take all the capacity — and a burst waits at its own gate instead of knocking everyone else over.

Shared pool: one burst, everyone refused TokenFusion: a fair share per account

A fair share per account

Each account has its own limit on calls in flight, checked before anything is held or charged.

Clear signals, not hung sockets

At a limit you get an immediate 429 or 503 with Retry-After, so your client knows exactly when to try again.

Headroom for everything else

Capacity is always kept back for sign-in, top-ups and your dashboard, even when calls are at their peak.

What you get

Everything a team needs to build on AI — and nothing it has to babysit.

One OpenAI-compatible API

Point your existing OpenAI SDK at one base URL. Chat, streaming and image models sit behind the same endpoint.

A key for each model

Create a key for any model in a click, with ready-to-paste code. Revoke one without touching the rest.

A prepaid balance in US dollars

Top up from $1. Your balance never expires, and it cannot go negative.

Metered to the millionth

Every call is charged at the published selling rate for its model, shown to six decimal places, with per-model totals.

Teams with real budgets

Fund each team from your wallet, cap each member by the month, and see who used what.

Know where it runs

Every model states the region it is served from. If we have not been told, we say so — we never guess.

Refused, never invented

If a model cannot be served, the call is refused and nothing is billed. You never get a made-up answer.

Receipts and a full history

A receipt for every top-up and a per-call usage history you can check line by line.

Pricing

See what a call really costs — before you make one.

Pick what you're doing and how often. Every figure comes from our live rate card, in US dollars, worked out exactly the way a call is billed. You are charged at the published selling rate shown for the model, on the usage reported for each call. If a call comes back with no usage, it is charged on an estimate — and marked as one on your usage page.

Up to 70%below list price

We buy capacity in volume. Where we pay less than list, we pass part of it on — the list price is shown crossed out beside your price below, so you can check it.

+8%more balance on a $500 top-up

Pay $500 and $540.00 lands in your balance, so every call on every model costs you less.

$0lost to expiry

Your balance never expires. Nothing you top up is wasted on a deadline.

Per tokenand nothing else

No subscription, no seats, no minimum term. Each call is charged on the tokens it reports, to the millionth of a dollar.

$0for a refused call

If a model cannot be served, the call is refused and nothing is billed.

1What are you doing?
2How many calls a month?
Lowest cost for quick chat deepseek-v3.2
Per call$0.000003
Per month · 10,000 calls $0.03
ModelInput /1MOutput /1M Cost per callPer month
deepseek-v3.2 · global-ec-11
DeepSeek · Premium
$0.0022 $0.0029
$0.000003 $0.000004−28%
$0.03saves $0.01
MiniMax-M2.1 · global-ec-07
MiniMax · Premium
$0.011 $0.0433
$0.000024 $0.000033−28%
$0.24saves $0.09
deepseek-v4-flash · global-ec-12
DeepSeek · Premium
$0.0977 $0.2447
$0.000171 $0.000238−28%
$1.71saves $0.67
gemini-3.1-flash-lite · global-ec-16
Google · Premium
$0.1837 $1.102
$0.000514 $0.000714−28%
$5.14saves $2.00
MiniMax-M2.5 · global-ec-08
MiniMax · Premium
$0.2571 $1.0286
$0.000566 $0.000786−28%
$5.66saves $2.20
kimi-k2.6 · global-ec-22
Moonshot · Premium
$0.3315 $1.3776
$0.000745 $0.002483−70%
$7.45saves $17.38
glm-5.1 · ap-southeast-ec-34
Zhipu · Premium
$0.7347 $2.9388
$0.001616 $0.002245−28%
$16.16saves $6.29
kimi-k2.7-code · global-ec-23
Moonshot · Premium
$0.7957 $3.3061
$0.001788 $0.002483−28%
$17.88saves $6.95
claude-haiku-4-5 · global-ec-10
Anthropic · Premium
$0.7347 $3.6735
$0.001837 $0.002551−28%
$18.37saves $7.14
glm-5.2 · global-ec-18
Zhipu · Optimized
$0.9771 $3.4281
$0.002006 $0.002786−28%
$20.06saves $7.80
deepseek-v4-pro · global-ec-13
DeepSeek · Premium
$1.4694 $2.9388
$0.002351 $0.003265−28%
$23.51saves $9.14
Kimi K2.6 · global-fw-49
Moonshot AI · Premium · 262k context
$1.14 $4.80
$0.00258
$25.80
qwen3.7-max · global-ec-33
Alibaba · Premium
$1.4694 $4.4082
$0.002792 $0.003878−28%
$27.92saves $10.86
Kimi K3 · global-fw-36
Moonshot AI · Premium · 1M context
$3.60 $18.00
$0.009
$90.00
qwen3.6-max-preview · global-ec-30
Alibaba · Premium
$11.0204 $66.1224
$0.030857 $0.042857−28%
$308.57saves $120.00

Image, video and speech models

Priced per image, per second of video or per 1,000 characters of speech, as billed — not by the tasks above.

dreamina-seedance-2-0-hc · global-ec-14Video generation · Other 720p $0.121959 ($0.195429 with audio) · 1080p $0.244653 ($0.318123 with audio) per second of video
dreamina-seedance-2-0-mini-hc · global-ec-15Video generation · Other 720p $0.121959 ($0.195429 with audio) · 1080p $0.244653 ($0.318123 with audio) per second of video
happyhorse-1.1-t2v · global-ec-20Video generation · Other 720p $0.110204 · 1080p $0.146939 · 480p $0.055102 per second of video
kling/kling-v3-omni-video-generation · global-ec-24Video generation · Other 720p $0.113877 ($0.121959 with audio) · 1080p $0.20351 ($0.211592 with audio) per second of video
openai/gpt-image-2 · global-ec-25Image generation · OpenAI $4.9959 in · $18.7347 out per 1M tokens
qwen-audio-3.0-tts-flash · global-ec-26Audio · Alibaba $0.012196 per 1,000 characters spoken
qwen-image-2.0 · global-ec-27Image generation · Alibaba $0.024245 per image · 1024x1024
qwen-image-3.0 · global-ec-28Image generation · Alibaba $0.022041 per image · 1024x1024

Each figure is what one call with exactly these token counts is charged on that model; your real charge follows the tokens each call reports. Per month is one call times the number of calls. A crossed-out figure is the same call at the model's list price, shown only where your price is lower.

Top-ups

Any whole-dollar amount from $1 to $10,000; bigger top-ups add a bonus to your balance. No contracts, no minimum term. Your balance stays on the account for as long as the account is open.

See the full rate card
  • Starter $25
  • Team $100 +$5.00 bonus
  • Scale $500 +$40.00 bonus

Stop routing around the problem.

Open an account in a minute, top up, and call the model you need with one key.