⚡ New — Kimi K3 is live: bring your own Moonshot key →

50 providers. One gateway.

Every provider below works through the same BharatRouter API — same SDKs, same INR billing, same routing and failover. Bring your own key (BYOK): save your key once (encrypted, never shown again) and everything that provider serves becomes routable — even models beyond our catalog, as provider/model.

Save your keysHow BYOK works

50 providers

Krutrim Cloud

🇮🇳 India residencybring your own key

Ola's Krutrim Cloud — India's own AI cloud, serving Krutrim models and open weights (gpt-oss, Gemma 4, Qwen) from Indian datacenters. The natural home for india_only routing.

Details →

Sarvam AI

🇮🇳 India residencybring your own key

India's sovereign-LLM flagship — sarvam-30b and sarvam-105b understand 11 Indian languages and are India-resident, available via BYOK.

Details →

OpenAI

globalbring your own key

GPT-5, GPT-5-mini and GPT-4o-mini — the global frontier with BharatRouter routing, INR billing and automatic failover.

Details →

OpenRouter

globalbring your own key

A meta-gateway that unlocks hundreds of global models — Claude, Gemini, DeepSeek and more — through one upstream.

Details →

Groq

globalbring your own key

LPU inference — some of the fastest tokens in the industry for Llama, Qwen, Kimi K2 and gpt-oss models.

Details →

Mistral AI

globalbring your own key

Europe's frontier lab — Mistral Large and Medium plus the open Mixtral and Ministral families.

Details →

DeepSeek

globalbring your own key

DeepSeek V3 and R1 — frontier-class open-weight chat and reasoning models at disruptive prices.

Details →

Together AI

globalbring your own key

Serverless inference and fine-tuning across 200+ open-weight models.

Details →

Fireworks AI

globalbring your own key

Low-latency serving of open models with strong function-calling and JSON-mode support.

Details →

Cerebras

globalbring your own key

Wafer-scale hardware serving open models at thousands of tokens per second.

Details →

xAI

globalbring your own key

Grok models from xAI.

Details →

Moonshot AI

globalbring your own key

Kimi K2 — the trillion-parameter open-weight MoE family from Moonshot AI.

Details →

DeepInfra

globalbring your own key

Pay-per-token hosting for the most popular open-weight models.

Details →

Novita AI

globalbring your own key

Budget GPU cloud with a broad open-model catalog.

Details →

SambaNova

globalbring your own key

RDU-accelerated inference serving open models at very high throughput.

Details →

Nebius AI Studio

globalbring your own key

AI Studio serving open models from European datacenters.

Details →

Hyperbolic

globalbring your own key

Open-model serving on a decentralized GPU marketplace.

Details →

NVIDIA NIM

globalbring your own key

NVIDIA NIM — optimized inference microservices for open models on NVIDIA infrastructure.

Details →

Perplexity

globalbring your own key

Sonar models with built-in web grounding for answer-style completions.

Details →

AI21 Labs

globalbring your own key

Jamba hybrid SSM-Transformer models with very long context windows.

Details →

Upstage

globalbring your own key

Solar models — compact, capable LLMs with strong document-AI roots.

Details →

MiniMax

globalbring your own key

MiniMax long-context models from one of China's leading labs.

Details →

Alibaba Qwen (Intl)

globalbring your own key

The Qwen family served from Alibaba Cloud's international regions.

Details →

Z.ai (Zhipu)

globalbring your own key

GLM models from Z.ai, including the open GLM-4 line.

Details →

Featherless

globalbring your own key

Serverless access to thousands of Hugging Face model checkpoints.

Details →

kluster.ai

globalbring your own key

Distributed inference cloud for open-weight models.

Details →

Lambda

globalbring your own key

Lambda's inference API on their GPU cloud.

Details →

Chutes

globalbring your own key

Decentralized serverless AI compute built on Bittensor.

Details →

Cohere

globalbring your own key

Command models built for enterprise RAG and tool use.

Details →

Google Gemini

globalbring your own key

Google's Gemini family through the AI Studio endpoint.

Details →

AI/ML API

globalbring your own key

One API for 300+ models across providers — aggregator access with usage-based pricing.

Details →

FriendliAI

globalbring your own key

GPU-efficient serving of open-weight models through fast, low-cost serverless endpoints.

Details →

Inception (Mercury)

globalbring your own key

Mercury — diffusion-based LLMs that generate text in parallel for very low latency.

Details →

Jina AI

globalbring your own key

Embeddings and rerankers (jina-embeddings, jina-reranker) behind an OpenAI-compatible API.

Details →

OVHcloud AI Endpoints

globalbring your own key

European sovereign-cloud inference for popular open-weight models.

Details →

Venice AI

globalbring your own key

Privacy-first inference — no prompt logging or retention, open-weight models only.

Details →

SiliconFlow

globalbring your own key

A large aggregator of open-weight models — Qwen, DeepSeek, GLM and more — at low prices.

Details →

Baseten

globalbring your own key

Production model serving with fast cold-starts via OpenAI-compatible Model APIs.

Details →

Nscale

globalbring your own key

Sustainable GPU cloud with serverless inference for open-weight models.

Details →

Meta Llama API

globalbring your own key

Meta's official Llama API — the latest Llama models straight from the source.

Details →

StepFun

globalbring your own key

The Step family of multimodal and reasoning models from StepFun.

Details →

GMI Cloud

globalbring your own key

GPU cloud with serverless inference for open-weight models.

Details →

Scaleway

globalbring your own key

European cloud's Generative APIs — open-weight models served from the EU.

Details →

DigitalOcean Gradient

globalbring your own key

DigitalOcean's Gradient inference platform for open-weight models.

Details →

Volcano Engine (ByteDance)

globalbring your own key

ByteDance's cloud — Doubao and open models via the Ark OpenAI-compatible API.

Details →

Baidu ERNIE

globalbring your own key

Baidu's ERNIE models via the Qianfan platform.

Details →

Parasail

globalbring your own key

On-demand serverless inference for open-weight models.

Details →

W&B Inference

globalbring your own key

Weights & Biases Inference — hosted open models alongside the W&B tooling.

Details →

ElevenLabs

globalbring your own key

Voice AI — high-fidelity text-to-speech and speech-to-text, available via BYOK for the audio plane.

Details →

GitHub Models

globalbring your own key

GitHub's hosted model catalog (OpenAI, Llama, Mistral and more) reached with a GitHub token.

Details →

Why one gateway for every provider

Indian teams end up holding keys for half a dozen AI providers — an Indian lab for Indic language work, a frontier lab for hard reasoning, a fast-inference cloud for latency-critical paths. Each comes with its own dashboard, billing currency, rate limits and failure modes. BharatRouter collapses that into one API key, one INR invoice and one routing policy: optimize each request for price,latency or uptime, pin a provider when you must, and let circuit-breaker failover handle the rest.

Residency is a first-class request field, not a promise: data_policy: "india_only" restricts routing to India-resident endpoints like Krutrim Cloud and Sarvam AI — which is why they lead this page. Your provider keys are encrypted with AES-256-GCM, decrypted only in-flight per request, and revocable in one click. Agents get the same surface programmatically: an MCP server, a machine-readable catalog and scoped keys they can mint themselves.

Frequently asked questions

What does "bring your own key" (BYOK) mean on BharatRouter?

You save a provider's API key once in your BharatRouter account — encrypted with AES-256-GCM, never shown again — and every request to that provider rides your own account and your negotiated rates. BharatRouter adds routing, failover, one bill-view and team sharing on top. BYOK requests are free during beta.

Do I need my own key for every provider?

Yes — BharatRouter routes on your own provider keys (BYOK). Save a key once for each provider you want to use and BharatRouter handles routing, failover and one bill-view across all of them. There is no platform key to fall back to.

Can I use models that are not in the BharatRouter catalog?

Yes. When you save a provider key, BharatRouter discovers every model that key can serve and makes them routable as provider/model — for your organisation only, beyond the public catalog.

Which providers serve from India?

Krutrim Cloud and Sarvam AI serve from Indian datacenters. Send data_policy: "india_only" and BharatRouter routes only to India-resident endpoints — DPDP-aligned by construction.

How do I add a provider key?

Sign in, open Account Settings → Provider keys (BYOK), pick the provider and paste the key. It is verified live against the provider, encrypted at rest, and active within a minute.

Get a keyBrowse models