◇ The OpenAI-compatible gateway

Every model.
One endpoint.
Infinite fallbacks.

A production-grade AI gateway that routes across multiple providers, optimises for cost & latency, and never goes down.

∞
Auto failover

If one provider fails, the next one picks up instantly. Zero downtime.

◇
OpenAI-compatible

Drop-in replacement. Change base_url and api_key — nothing else.

◎
Multi-provider

OpenAI, Anthropic, DeepSeek, Groq, Gemini, Ollama, and any custom endpoint.

≡
Full request logs

Every request logged with tokens, latency, cost, and provider used.

◓
Cost analytics

Track spend per model, per provider, per user — in real time.

✦
Tiered access

v1 for all users. Upgrade to v2/v3 for more models and higher limits.

◇ Three versions, one gateway
/v1
Standard

Default tier — all users. Full OpenAI compatibility. Free-tier models included.

base_url = https://saki-gateway.indevs.in/v1
/v2
Pro

Higher rate limits, priority routing, access to premium models (GPT-4, Claude Sonnet, etc).

base_url = https://saki-gateway.indevs.in/v2
/v3
Enterprise

Dedicated capacity, custom providers, SLA-backed uptime, full audit logs.

base_url = https://saki-gateway.indevs.in/v3
◇ Quick start
cURL
curl -X POST https://saki-gateway.indevs.in/v1/chat/completions \
  -H "Authorization: Bearer sk-your-key" \
  -H "Content-Type: application/json" \
  -d '{"model":"auto","messages":[{"role":"user","content":"hello"}]}'
Python (openai SDK)
from openai import OpenAI
client = OpenAI(
    base_url="https://saki-gateway.indevs.in/v1",
    api_key="sk-your-key",
)
resp = client.chat.completions.create(
    model="auto",
    messages=[{"role": "user", "content": "hello"}],
)
print(resp.choices[0].message.content)