Access Claude, GPT, Gemini and more through a single OpenAI-compatible endpoint. No VPN needed. Pay only for what you use.
// One endpoint, every workflow //
The same requests you already write, routed to whichever model fits the job — no provider accounts, no regional workarounds.
TIPMix models inside one workflow — plan with Opus, execute with Sonnet.
No account juggling, no regional workarounds. One key, one endpoint, and the same request shape you already write.
Drop-in replacement. Change one line of code — your base URL — and access any supported model instantly.
Reliable connectivity worldwide. No VPN or proxy configuration required. Just works, everywhere.
Optimized routing with streaming support. Average response time under 3 seconds for most models.
No subscriptions, no commitments. Top up your balance and pay only for the tokens you consume.
Multi-channel load balancing with automatic failover. Built for consistent, reliable access.
Real-time usage dashboard. Track every request, monitor spending, and manage API keys with ease.
Every card shows the list price struck through, then your KissAPI rate. Same weights, same context window, same streaming API.
First when you top up, again when a request is billed. The struck-through figures on the cards above are vendor list prices — here is how far below them you actually land.
Claude Opus 5 output, per 1M tokens:$25.00 list → $20.00 in credits → $4.00 out of your pocket.
All prices in USD. No hidden fees. Pay per token, top up anytime.
Blended estimate at a 3:1 input/output ratio. Actual usage depends on your prompts.
// Use cases //
One endpoint covers the whole stack — editors, agents, batch jobs and production traffic.
Use Claude and GPT-5 in Cursor, Windsurf, Cline, or any editor that speaks the OpenAI format. One key, every model.
Run Claude Code, Codex CLI or your own framework. Multi-channel routing fails over without touching your code.
Classify documents, generate embeddings, run evals at scale. Per-token billing keeps large jobs predictable.
Ship AI features without juggling provider accounts. Unified billing, usage logs and model switching in one place.
Point Cherry Studio, Open WebUI or LobeChat at one endpoint and switch models mid-conversation.
Benchmark Opus against Sonnet and GPT-5 side by side on identical prompts, then compare cost per run.
From the Blog