Bills nobody can explain
Provider invoices show totals, not reasons. InstaRoute shows the cost of every request next to what a premium model would have charged.
InstaRoute is the cost-control layer between your apps and 29 AI providers. It judges how hard each request is, sends it to the cheapest model that will do the job well, detours around failing providers and itemises what you saved — through one OpenAI- and Anthropic-compatible API.
Most teams wire every feature to one premium model, through one provider, and find out what it cost at month-end. InstaRoute replaces that single wire with a routing layer that knows what each request needs — and keeps the receipts.
Provider invoices show totals, not reasons. InstaRoute shows the cost of every request next to what a premium model would have charged.
A 429 or an expired key stops your product. InstaRoute detours to an equivalent model on another provider automatically.
One OpenAI- and Anthropic-compatible endpoint covers 29 providers. Change models with a setting, not a release.
Emails, phone numbers, Aadhaar and PAN are masked before any provider sees them — and put back in the answer.
Every request takes the same short drive through InstaRoute — usually a few milliseconds — and leaves on the lane its difficulty deserves. This is how the LLM router decides.
Personal data and secrets are masked; prompt injection is flagged or blocked.
An identical or near-identical question is answered from cache, for ₹0.
A difficulty score decides how much model the request really needs.
Express, Standard or Frontier — ranked on cost, speed and quality, with detours on failure.
The response carries the model, cost, baseline and saving; Logs keep the rest.
Point any OpenAI or Anthropic SDK at InstaRoute and keep shipping. Streaming, tools, vision and JSON mode behave exactly as before.
Connect the provider keys you already pay for. InstaRoute never buys or resells tokens, so there is nothing to mark up.
Each prompt is scored for difficulty in milliseconds — easy work goes to fast, low-cost models, hard work to frontier ones.
A 429, an outage, an expired key or an empty credit balance triggers an instant detour to an equivalent model elsewhere.
Repeat and near-repeat questions are answered from cache. Prompt-cache-aware routing and a flex tier trim the rest.
Emails, phone numbers, cards, Aadhaar and PAN are masked before a provider sees them — and restored in the reply.
Model, attempts, tokens, cost, premium baseline and saving — in the response, in Logs and in your monthly statement.
Budgets per key, a ceiling per workspace, and Slack or webhook alerts long before anyone gets a surprise.
Import your own benchmark scores and shadow-test a candidate model on a slice of real traffic.
Add a vLLM or any OpenAI-compatible endpoint, or Ollama for fully local inference, and route to it like any provider.
14 embedding models from 7 providers behind the same key and a single /v1/embeddings endpoint.
Owner, admin, member and viewer roles, invite links, per-app keys and an audit log of every change.
InstaRoute is an OpenAI-compatible API that also speaks Anthropic. Change the base URL, use an sk-inf-… key, and ask for a goal — auto, cheapest, fastest, coding — or any model by name. Every reply brings its receipt in an instaroute field.
Set one environment variable and every coding session gets failover between Claude providers, per-developer budgets, cost and savings in one dashboard, and warm prompt caches — without sharing your company's Anthropic key.
# Point Claude Code at InstaRoute $ export ANTHROPIC_BASE_URL=https://instaroute.instadatahelp.com $ export ANTHROPIC_AUTH_TOKEN=sk-inf-… $ claude ✻ Welcome to Claude Code /status → base URL instaroute.instadatahelp.com ✓
A quick way to see how much routing could reduce your LLM costs. Once you are live, every rupee is measured per request — and the 3% only applies to what is measured.
An estimate to help you decide. Your real savings are measured on every request — the fee is based on those measurements only.
Each request is logged with the model that answered, what it cost, what the premium model would have cost, and the difference. Prompt text is never stored.
Screens shown with demo data.
Compared with Inferect, OpenRouter, Portkey and LiteLLM, using each vendor's own website.
| Capability | InstaRoute | Inferect | OpenRouter | Portkey | LiteLLM |
|---|---|---|---|---|---|
| Starting price | ₹1,000 + GST / $20 a month, all features | $17 / ₹1,199 (launch offer) | No plan fee | $49 a month + usage | Free (self-host) |
| Fee on spend or savings | 3% of measured savings | 2–3% of savings | 5.5% on credits | None | None |
| Your own provider keys, no markup | ✔ | ✔ | ✔ | ✔ | ✔ |
| Automatic per-request model choice | ✔ | ✔ | ✔ | partial | ✔ |
| Claude Code & Anthropic API | ✔ | partial | ✔ | ✔ | ✔ |
| Semantic caching | ✔ | partial | — | ✔ | ✔ |
| PII redaction with Aadhaar & PAN | ✔ | — | — | — | — |
| Savings vs baseline on every request | ✔ | partial | — | — | — |
| INR billing, UPI & GST tax invoice | ✔ | partial | — | — | — |
✔ available · partial = limited or plan-dependent · — not stated on the vendor's website. Based on each vendor's public website and documentation, checked October 2026; prices exclude taxes and change often. Product names are trademarks of their respective owners.
or US$20 / month outside India
+ 3% of the savings we measure in the previous month
Start your free trial →Pay by UPI, cards, net banking or wallets (PayU) · cards via Stripe outside India · GST tax invoice with your GSTIN
An AI gateway sits between your application and AI providers such as OpenAI, Anthropic and Google. An LLM router decides which model should answer each request. InstaRoute does both: one OpenAI-compatible API in front of 29 providers that picks the best-value model for every request, fails over automatically and records the cost and savings.
Most requests do not need the most expensive model. InstaRoute scores each request for difficulty and sends simple ones to fast, low-cost models and only hard ones to premium models. Exact and semantic caching, prompt-cache-aware routing and a flex tier add further savings. Every response shows the cost against the premium baseline, so the savings are measured, not estimated.
Yes. Change the base URL to https://instaroute.instadatahelp.com/v1 and use your InstaRoute key — the OpenAI SDKs for Python, Node, .NET, Go and Java, LangChain, LlamaIndex and the Vercel AI SDK all work. Streaming, tool calling, vision and JSON mode are supported.
Yes. InstaRoute speaks the Anthropic Messages API. Set ANTHROPIC_BASE_URL to https://instaroute.instadatahelp.com and ANTHROPIC_AUTH_TOKEN to your InstaRoute key, then run claude. Each session gets failover, cost tracking and per-key budgets.
No. You bring your own provider keys and pay providers directly at their prices — InstaRoute never marks up tokens. You pay InstaRoute ₹1,000 + GST a month (US$20 outside India) plus 3% of the savings it measures.
29 providers including OpenAI, Anthropic, Google Gemini, Azure OpenAI, xAI Grok, Mistral, DeepSeek, Groq, Cerebras, Together, Fireworks, OpenRouter, Perplexity, Cohere, Alibaba Qwen, Moonshot Kimi and Z.ai GLM, plus Ollama and any OpenAI-compatible endpoint you host — over 170 routable models.
Prompt and reply text are never written to logs or the database — only metadata such as model, tokens and cost. Provider keys are encrypted at rest (AES-256-GCM), and guardrails can mask emails, phone numbers, card, Aadhaar and PAN numbers before any provider sees them.
After the 7-day free trial you get one bill every 30 days: ₹1,000 plus 3% of the previous period's measured savings, with 18% GST shown separately. Pay by UPI, card or net banking through PayU and receive a GST tax invoice with your GSTIN instantly. There is no auto-debit and no contract.
Connect one provider key, change one base URL, send one request. Every feature, free for 7 days — no card needed.
instaroute.instadatahelp.com · a product of InstaDataHelp AI Services Private Limited