Anthropic
3 models
- large
Claude Fable 5
Frontier model, autonomous coding
- medium
Claude Sonnet 4.6
Balanced quality and cost
- small
Claude Haiku 4.5
Ultra-fast, lightweight tasks
One endpoint · every frontier model
Point your tools at codai. The router picks the right frontier model for each task, checks that it actually worked, and you only pay for verified successes.
Drop-in for any OpenAI-compatible client
# any OpenAI-compatible clientnpm i @codai/sdkbase_url: https://ai.codai.ro/v1model: codai
Same request shape, one model name. No SDK rewrite.
Works in
How it works
You keep the client you already use. codai does the routing, the verification and the accounting.
Set the base URL to ai.codai.ro/v1 and the model to codai. Copilot, Cursor, Claude Code, Codex CLI or any OpenAI SDK — nothing else changes.
Every request is classified and sent to the model that wins on that kind of task — Claude, GPT or Gemini — with automatic failover when a provider degrades.
A task is only billed when it is verified as successful — by tests, by the harness, or by your one-tap confirmation. Failed attempts cost you nothing.
What you get
Everything an agentic workflow needs, behind one name.
Per-task model selection across Claude, GPT and Gemini, driven by open benchmarks and live production telemetry.
HTTP, Python, shell and browser automation run server-side in sandboxes — so any client can use them, even without native tool support.
Facts, decisions and repo conventions survive across sessions and devices. Scoped per user, exportable, deletable.
Attach your own Model Context Protocol servers once and every codai client gets the same tools.
Outcome billing: a flat charge per verified successful task instead of a token meter you cannot predict.
Use your existing Anthropic, OpenAI or Google keys through the same router and tooling, from €5 a month.
Hosted in europe-west1 on Google Cloud Run. Prompts and usage stay in the EU; no training on your data by default.
Routing decisions are backed by public, reproducible benchmark runs. Check the numbers before you trust the router.
Model roster
codai selects across Claude, GPT and Gemini. When a new model ships, we benchmark it and add it to the pool — you never rename anything.
3 models
Claude Fable 5
Frontier model, autonomous coding
Claude Sonnet 4.6
Balanced quality and cost
Claude Haiku 4.5
Ultra-fast, lightweight tasks
3 models
GPT-5
Frontier reasoning model
GPT-5 mini
Fast, cost-effective tasks
GPT-5 nano
High-throughput micro-tasks
3 models
Gemini 2.5 Pro
Long context, multimodal
Gemini Flash 2.5
Low latency, high throughput
Gemini Flash 2.0
Sub-10 ms, experimental
Use model="codai" in any OpenAI-compatible client. codai handles the rest.
Benchmarks
Every routing decision is grounded in public benchmark runs over real task suites — coding, agentic tool use, long context. Same harness for every model, results published as they land.
Pricing
Every plan includes a monthly number of verified successful tasks. Beyond that, a flat €1.20 per verified success — never a token meter.
Try the router with real tasks.
Free
5 tasks included
€1.20 per successful task beyond your plan
For one developer shipping every week.
€40/month
33 tasks included
€1.20 per successful task beyond your plan
Cancel anytime
For builders running agents daily.
€100/month
83 tasks included
€1.20 per successful task beyond your plan
Cancel anytime
For teams with production workloads.
€300/month
250 tasks included
€1.20 per successful task beyond your plan
Cancel anytime
Have your own provider keys? BYOK from €5/month with a 7-day trial.
See all plansTrust
Your prompts, keys and usage data are handled the way you would handle them yourself.
Gateway, database and memory run on Google Cloud Run in europe-west1 (Belgium). Nothing is stored outside the EU.
Prompts and completions are never used to train models unless you explicitly opt in from your account.
Export everything we hold about you, or delete your account and data, from your account page — no support ticket needed.
Sign in with passkeys; mint API keys with per-key budgets and model scopes, and revoke them instantly.
No card needed. Point your client at codai and see where the router sends your first request.
Live routing
Real production traffic — the actual distribution of upstream models the codai router selected. Updated every five minutes.
13,569 routed requests · last 24 hours · successful codai requests only