Esc to close · ⌘K / Ctrl-K opens search anywhere
Every provider below works through the same BharatRouter API — same SDKs, same INR billing, same routing and failover. Bring your own key (BYOK): save your key once (encrypted, never shown again) and everything that provider serves becomes routable — even models beyond our catalog, as provider/model.
50 providers
🇮🇳 India residencybring your own key
Ola's Krutrim Cloud — India's own AI cloud, serving Krutrim models and open weights (gpt-oss, Gemma 4, Qwen) from Indian datacenters. The natural home for india_only routing.
🇮🇳 India residencybring your own key
India's sovereign-LLM flagship — sarvam-30b and sarvam-105b understand 11 Indian languages and are India-resident, available via BYOK.
globalbring your own key
GPT-5, GPT-5-mini and GPT-4o-mini — the global frontier with BharatRouter routing, INR billing and automatic failover.
globalbring your own key
A meta-gateway that unlocks hundreds of global models — Claude, Gemini, DeepSeek and more — through one upstream.
globalbring your own key
LPU inference — some of the fastest tokens in the industry for Llama, Qwen, Kimi K2 and gpt-oss models.
globalbring your own key
Europe's frontier lab — Mistral Large and Medium plus the open Mixtral and Ministral families.
globalbring your own key
DeepSeek V3 and R1 — frontier-class open-weight chat and reasoning models at disruptive prices.
globalbring your own key
Serverless inference and fine-tuning across 200+ open-weight models.
globalbring your own key
Low-latency serving of open models with strong function-calling and JSON-mode support.
globalbring your own key
Wafer-scale hardware serving open models at thousands of tokens per second.
globalbring your own key
Kimi K2 — the trillion-parameter open-weight MoE family from Moonshot AI.
globalbring your own key
RDU-accelerated inference serving open models at very high throughput.
globalbring your own key
NVIDIA NIM — optimized inference microservices for open models on NVIDIA infrastructure.
globalbring your own key
Sonar models with built-in web grounding for answer-style completions.
globalbring your own key
Jamba hybrid SSM-Transformer models with very long context windows.
globalbring your own key
The Qwen family served from Alibaba Cloud's international regions.
globalbring your own key
Serverless access to thousands of Hugging Face model checkpoints.
globalbring your own key
One API for 300+ models across providers — aggregator access with usage-based pricing.
globalbring your own key
GPU-efficient serving of open-weight models through fast, low-cost serverless endpoints.
globalbring your own key
Mercury — diffusion-based LLMs that generate text in parallel for very low latency.
globalbring your own key
Embeddings and rerankers (jina-embeddings, jina-reranker) behind an OpenAI-compatible API.
globalbring your own key
European sovereign-cloud inference for popular open-weight models.
globalbring your own key
Privacy-first inference — no prompt logging or retention, open-weight models only.
globalbring your own key
A large aggregator of open-weight models — Qwen, DeepSeek, GLM and more — at low prices.
globalbring your own key
Production model serving with fast cold-starts via OpenAI-compatible Model APIs.
globalbring your own key
Sustainable GPU cloud with serverless inference for open-weight models.
globalbring your own key
Meta's official Llama API — the latest Llama models straight from the source.
globalbring your own key
European cloud's Generative APIs — open-weight models served from the EU.
globalbring your own key
DigitalOcean's Gradient inference platform for open-weight models.
globalbring your own key
ByteDance's cloud — Doubao and open models via the Ark OpenAI-compatible API.
globalbring your own key
Weights & Biases Inference — hosted open models alongside the W&B tooling.
globalbring your own key
Voice AI — high-fidelity text-to-speech and speech-to-text, available via BYOK for the audio plane.
globalbring your own key
GitHub's hosted model catalog (OpenAI, Llama, Mistral and more) reached with a GitHub token.
Indian teams end up holding keys for half a dozen AI providers — an Indian lab for Indic language work, a frontier lab for hard reasoning, a fast-inference cloud for latency-critical paths. Each comes with its own dashboard, billing currency, rate limits and failure modes. BharatRouter collapses that into one API key, one INR invoice and one routing policy: optimize each request for price,latency or uptime, pin a provider when you must, and let circuit-breaker failover handle the rest.
Residency is a first-class request field, not a promise: data_policy: "india_only" restricts routing to India-resident endpoints like Krutrim Cloud and Sarvam AI — which is why they lead this page. Your provider keys are encrypted with AES-256-GCM, decrypted only in-flight per request, and revocable in one click. Agents get the same surface programmatically: an MCP server, a machine-readable catalog and scoped keys they can mint themselves.
You save a provider's API key once in your BharatRouter account — encrypted with AES-256-GCM, never shown again — and every request to that provider rides your own account and your negotiated rates. BharatRouter adds routing, failover, one bill-view and team sharing on top. BYOK requests are free during beta.
Yes — BharatRouter routes on your own provider keys (BYOK). Save a key once for each provider you want to use and BharatRouter handles routing, failover and one bill-view across all of them. There is no platform key to fall back to.
Yes. When you save a provider key, BharatRouter discovers every model that key can serve and makes them routable as provider/model — for your organisation only, beyond the public catalog.
Krutrim Cloud and Sarvam AI serve from Indian datacenters. Send data_policy: "india_only" and BharatRouter routes only to India-resident endpoints — DPDP-aligned by construction.
Sign in, open Account Settings → Provider keys (BYOK), pick the provider and paste the key. It is verified live against the provider, encrypted at rest, and active within a minute.