Skip to content
Private beta · onboarding early teams
Agent-Native AI Gateway

One key. Every model.
Built for AI agents.

One OpenAI-compatible key, every major model behind it — with routing, failover, caching, and credit accounting in one layer.

Read the docs →

Already in a coding agent? Tell it:

https://aisetu.ai/llms.txt Integrate AI Setu
Coding Agentbuilds & shipsEditor Agentin your IDEAnalytics Agentqueries dataOpenAIproviderAnthropicproviderGeminiproviderBedrockproviderAzureproviderAI Setu/v1 · routing

6 providers natively, any OpenAI-compatible endpoint besides

OpenAIAnthropicGoogle GeminiAWS BedrockAzure OpenAIVertex AIGroqvia OpenAI-compatiblexAIvia OpenAI-compatibleDeepSeekvia OpenAI-compatibleMistralvia OpenAI-compatiblevLLM & Ollamavia OpenAI-compatibleOpenAIAnthropicGoogle GeminiAWS BedrockAzure OpenAIVertex AIGroqvia OpenAI-compatiblexAIvia OpenAI-compatibleDeepSeekvia OpenAI-compatibleMistralvia OpenAI-compatiblevLLM & Ollamavia OpenAI-compatible
Under the hood

What happens inside one request.

Authenticate, check the cache, route to the best provider, meter the spend — all in ~80µs of overhead, under 1% of a typical LLM call.

AI Setu · /v1 gatewayresponse · credits meteredYour agentone requestProviderbest of 81Authenticatekey · tenant2Cacheexact + semantic3Routeyour keys first4Metercredits + caps
Agent native

The gateway your agents can run themselves.

Every surface is an SDK or MCP server. The agents building your product sign up, get a key, and wire themselves in — no human glue.

# the agent onboards itself
claude mcp add ai-setu

 signed up via OTP
 workspace created · key issued
 your keys attached · anthropic, openai
# ready — agent is now wired
Features

Infrastructure that stays out of the way.

Routing, reliability, caching, and accounting in one layer. One integration, not five.

Model-based routing

Send a model name; we pick the provider — your keys first.

Failover before first byte

Outages reroute before the first token. No double billing, no partial output.

Your keys, your rates

Use your own provider accounts — pay providers directly, no markup. Keys encrypted at rest with AES-GCM; only a last-four ever returns.

Exact + semantic cache

Repeat and reworded calls serve from cache — a cache hit costs nothing at the provider.

Credit accounting

Metered in credits with live balance and daily caps. No surprise bills.

Multi-tenant & reseller-ready

Per-tenant isolation, audit logs, and one-call child-tenant provisioning.

Guardrails built in

PII redaction, prompt-injection detection, and content policy — inline, per workspace, before and after the model.

Usage & cost analytics

Spend sliced by provider, model, key, tag, cache, and prompt — one dashboard, no wiring.

Prompt Studio

Server-side templates, versioning, labels, a playground, and A/B tests — with cost attributed per prompt.

Regional edge

Built for India & MEA, not just US-West.

Multilingual routing across local and global providers, data-residency options, and regional failover — quality in your users' languages, not just English.

हिन्दीالعربيةதமிழ்తెలుగుBahasa+ English
Pricing

Pay for the routing, not the markup.

A flat fee per 1,000 routed requests — same whether you bring your keys or use ours. Cache hits are free.

Developer
$0 / mo

For trying it out and side projects.

  • 1 workspace, your own keys
  • Usage dashboard
  • Exact-match cache
  • Community support
Pro POPULAR
$49 / mo + usage

For teams shipping production agents.

  • Unlimited workspaces
  • Semantic cache + failover
  • Daily caps & rate limits
  • MCP server access
  • SSO + priority routing
Enterprise
Custom

Regulated, multilingual, platforms.

  • Data-residency options
  • Self-host substrate
  • SLA + dedicated support
  • India/MEA multilingual routing
  • Reseller child-tenant provisioning
$0.20 / 1,000 routed requests
$0.00 on cache hits
Zero markup on your own keys

Figures are indicative and pending Finance sign-off — placeholders to show structure, not final list prices.

Stop wiring providers

Start shipping agents.

One key, every provider, and a bridge your agents can build on themselves.