One price per model.
Always the fastest node.

Every model below routes across multiple providers. You pay the published rate no matter which one serves you, no platform fee.

NewAvailable

Kimi-K3

Moonshot · 1M context
82 tok/s0.75s TTFT
$3.00 / $15.00 per 1M
cached input $0.30 / 1M
NewAvailable

DeepSeek-V4-Flash

DeepSeek · 1M context
103 tok/s0.88s TTFT
$0.14 / $0.28 per 1M
cached input $0.04 / 1M
Available

DeepSeek-V4-Pro

DeepSeek · 1M context
102 tok/s0.75s TTFT
$1.74 / $3.48 per 1M
cached input $0.43 / 1M
Available

GLM-5.2

Z.ai · 1M context
161 tok/s0.83s TTFT
$1.40 / $4.40 per 1M
cached input $0.35 / 1M
Available

Kimi-K2.6

Moonshot · 256K context
157 tok/s0.83s TTFT
$0.95 / $4.00 per 1M
cached input $0.24 / 1M
Available

Kimi-K2.7-Code

Moonshot · 256K context
172 tok/s0.61s TTFT
$0.95 / $4.00 per 1M
cached input $0.24 / 1M
Available

MiMo-V2.5-Pro

Xiaomi · 1M context
137 tok/s1.49s TTFT
$0.80 / $3.00 per 1M
cached input $0.20 / 1M
Available

MiniMax-M3

MiniMax · 1M context
101 tok/s0.92s TTFT
$0.30 / $1.20 per 1M
cached input $0.07 / 1M
Available

Step-3.7-Flash

StepFun · 256K context
172 tok/s2.33s TTFT
$0.25 / $1.30 per 1M
cached input $0.06 / 1M
Available

DeepSeek-V3.2

DeepSeek · 160K context
25 tok/s1.40s TTFT
$0.50 / $1.50 per 1M
cached input $0.16 / 1M
Available

gemma-4-26b-a4b-it

Google · 1M context
153 tok/s1.03s TTFT
$0.13 / $0.40 per 1M
cached input $0.13 / 1M
Available

gemma-4-31b-it

Google · 1M context
94 tok/s1.00s TTFT
$0.15 / $0.46 per 1M
cached input $0.15 / 1M
Available

GLM-5

Z.ai · 200K context
61 tok/s1.22s TTFT
$1.00 / $3.20 per 1M
cached input $0.25 / 1M
Available

gpt-oss-120b

OpenAI · 128K context
162 tok/s0.41s TTFT
$0.15 / $0.60 per 1M
cached input $0.15 / 1M
Available

gpt-oss-20b

OpenAI · 128K context
173 tok/s0.25s TTFT
$0.07 / $0.30 per 1M
cached input $0.07 / 1M
Available

MiMo-V2.5

Xiaomi · 1M context
99 tok/s0.73s TTFT
$0.14 / $0.28 per 1M
cached input $0.05 / 1M
Available

MiniMax-M2.7

MiniMax · 192K context
86 tok/s1.25s TTFT
$0.30 / $1.20 per 1M
cached input $0.30 / 1M
Available

Mistral-Nemo-Instruct-2407

Mistral · 128K context
50 tok/s0.91s TTFT
$0.04 / $0.17 per 1M
cached input $0.04 / 1M
Available

Nemotron-3-Ultra-550B-A55B

NVIDIA · 500K context
195 tok/s0.59s TTFT
$0.60 / $2.40 per 1M
cached input $0.24 / 1M
Available

Qwen3.6-27B

Alibaba · 256K context
54 tok/s1.06s TTFT
$0.60 / $3.60 per 1M
cached input $0.60 / 1M
Available

Qwen3.6-35B-A3B

Alibaba · 256K context
186 tok/s0.70s TTFT
$0.25 / $1.49 per 1M
cached input $0.25 / 1M

Throughput and TTFT are live medians, measured continuously.

Predictable pricing

Top up, spend against it, auto-refill when low.

No commitment

Just plug in and start building. Stop anytime.

Volume & enterprise

High-volume discounts, dedicated support and SLAs at scale. Talk to us ↗

FAQ

How models are priced and billed on Keln. Still have questions?
Reach out anytime.

FAQ page ↗

One published rate per model, billed per token, the same rate no matter which provider or node serves the request. Keln takes no platform fee and adds no per-provider markup, so your cost is predictable even as routing changes underneath you.

No. Price is fixed per model, so Keln is free to send every request to the fastest verified capacity. Faster nodes cost you the same as slow ones.

Yes. Every model lists an input and an output rate per 1M tokens. Reasoning tokens are billed as output tokens, and cached input is billed at a reduced rate.

Prepaid credits. Top up a balance, spend against it, and set auto-refill so you never hit zero. Spend is visible live per project and per key. Keln auto-generates invoices for business.

You are billed once, for the tokens you receive. Failed attempts and Keln-initiated failovers are on us. Recovery is part of the service.

Contact us ↗ to get more information on high-volume discounts.

Ready to try Keln?

Start building in minutes.