One price per model.
Always the fastest node.
Every model below routes across multiple providers. You pay the published rate no matter which one serves you, no platform fee.
qwen3.8-2.4t-a95b
deepseek-v4-flash-0731
kimi-k3
deepseek-v4-pro
glm-5.2
kimi-k2.6
kimi-k2.7-code
mimo-v2.5-pro
minimax-m3
step-3.7-flash
gemma-4-26b-a4b-it
gemma-4-31b-it
gpt-oss-120b
gpt-oss-20b
mimo-v2.5
nemotron-3-ultra-550b-a55b
qwen3.6-27b
qwen3.6-35b-a3b
Throughput and TTFT are live medians, measured continuously.
Top up, spend against it, auto-refill when low.
Just plug in and start building. Stop anytime.
High-volume discounts, dedicated support and SLAs at scale. Talk to us ↗
One published rate per model, billed per token, the same rate no matter which provider or node serves the request. Keln takes no platform fee and adds no per-provider markup, so your cost is predictable even as routing changes underneath you.
No. Price is fixed per model, so Keln is free to send every request to the fastest verified capacity. Faster nodes cost you the same as slow ones.
Yes. Every model lists an input and an output rate per 1M tokens. Reasoning tokens are billed as output tokens, and cached input is billed at a reduced rate.
Prepaid credits. Top up a balance, spend against it, and set auto-refill so you never hit zero. Spend is visible live per project and per key. Keln auto-generates invoices for business.
You are billed once, for the tokens you receive. Failed attempts and Keln-initiated failovers are on us. Recovery is part of the service.
Contact us ↗ to get more information on high-volume discounts.