Simple, Predictable Plans for Every AI Workload
Orchestrate across Claude, Gemini, ChatGPT, and private local LLMs. Zero vendor lock-in, autonomous failovers, and transparent BYOK pricing.
Basic
For developers and independent builders starting with multi-LLM routing.
Billed annually (20% off)
Included capabilities:
- Unlimited Bring Your Own Keys (BYOK)
- Local LLM bridge (Ollama, vLLM)
- Basic latency & cost routing optimizer
- Side-by-side prompt test runner
- Single developer workspace
- 14-day execution trace & log history
- UPI, Netbanking & Card payments
Pro
For serious engineers and startups deploying production AI agents.
Billed annually (20% off)
Included capabilities:
- Everything in Basic, plus:
- Autonomous cost & latency routing optimizer
- Automated fallback cascades (zero downtime)
- LLM output evaluator & consensus voting
- Ultra-low latency edge gateway (<40ms overhead)
- 30-day analytics & granular token metrics
- Custom temperature & system prompt injection
- GST input tax credit invoice available
Scale / Team
For fast-growing product teams scaling multi-agent AI pipelines.
Billed annually (20% off)
Included capabilities:
- Everything in Pro, plus:
- 10 team seats included (additional at ₹799/seat)
- Multi-tenant API keys & role-based access (RBAC)
- Custom routing rules & weighted traffic split
- Automated prompt regression testing suite
- Real-time token cost allocation per customer/tenant
- 99.95% uptime service level agreement
- India & global edge point-of-presence
Enterprise
Dedicated infrastructure, security isolation, and bespoke LLM governance.
Billed annually (20% off)
Included capabilities:
- Everything in Scale / Team, plus:
- Dedicated router instances & VPC peering (AWS Mumbai, GCP Delhi, Azure India)
- Zero-Data Retention (ZDR) guarantee & audit trail
- Custom internal & on-prem fine-tuned model connectors
- ISO 27001, SOC 2 Type II & DPDP Act compliance package
- Custom Indian corporate procurement, PO & annual MSA
- 99.99% guaranteed uptime SLA
Calculate Your Token Savings
See how much your engineering team saves by routing tasks dynamically instead of sending every prompt to expensive frontier models.
40% deep reasoning, 40% fast structured extraction, 20% cached/local.
That equals approximately ₹85,200 reinvested into engineering per year.
No migration required · Works with existing OpenAI / Anthropic SDKs
Compare All Plan Features
Explore granular specifications, rate limits, enterprise controls, and support tiers.
| Features & Capabilities | Basic | Pro | Scale / Team | Enterprise |
|---|---|---|---|---|
Orchestration & Routing | ||||
Dynamic Multi-Model Router Intelligently routes prompts to the best model based on latency, cost, and task capability. | Standard Router | Autonomous AI Router | Autonomous AI Router | Custom Trained Policy Router |
Automatic Fallback Cascades Zero-error failovers if an upstream provider experiences downtime or rate limits. | ||||
Parallel Consensus Voting Broadcast prompts to multiple models concurrently and synthesize the highest-scored answer. | Up to 3 models | Up to 8 models | Unlimited models | |
Bring Your Own Keys (BYOK) Plug in your own OpenAI, Anthropic, Google, and Mistral API keys without commission. | ||||
Local LLM Bridge (Ollama / vLLM) Route requests seamlessly between private local hardware and cloud models. | ||||
Routing Overhead Latency Gateway proxy time added to raw upstream model latency. | <80ms | <40ms | <20ms | <5ms (Dedicated Edge) |
Model Access & Limits | ||||
Included Router Tokens Managed token quota included in base monthly subscription. | 100K / mo | 1.5M / mo | 6M / mo | Custom Pool |
Concurrent In-flight Requests Simultaneous open model connections. | 5 | 25 | 100 | Custom Dedicated |
Supported Foundation Models Access to latest Claude, Gemini, GPT-4o, DeepSeek, Llama models out of the box. | 15+ Models | 50+ Models | 50+ Models | All Models + Custom Models |
Custom Fine-tuned Endpoints Connect proprietary internal fine-tunes or HuggingFace endpoints. | ||||
Analytics, Tracing & Quality | ||||
Execution History & Tracing Inspection window for request payloads, prompts, tokens, and latency traces. | 14 days | 30 days | 90 days | Unlimited / S3 Archival |
Cost Allocation per User/Tenant Attribute token spend and cost metrics directly to specific end-users or API keys. | ||||
Prompt Regression Testing Suite Run golden dataset benchmarks before updating system prompts or model versions. | Manual Runner | Automated CI/CD | Full Golden Suite + Webhooks | |
OpenTelemetry Export Stream traces directly to Datadog, New Relic, or Prometheus. | ||||
Security, Governance & Support | ||||
Data Privacy & Zero Data Retention Strict policy guaranteeing customer prompt data is never stored or used to train models. | Standard Privacy | Standard Privacy | ZDR Available | Guaranteed Zero Retention + BAA |
Team Seats Collaborators and dashboard accounts. | 1 user | 1 user | 10 seats included | Unlimited |
Compliance & Invoicing Verified security certifications, GST compliance and auditor access. | GST Invoicing | GST Invoicing | GST + DPDP Compliant | ISO 27001, SOC 2, HIPAA, DPDP |
Deployment Topology Where the orchestration routing proxy runs. | Multi-tenant Cloud | Multi-tenant Cloud | Priority Multi-tenant | Dedicated VPC (India / Global) |
Support SLA Response time and support availability. | Email & Discord | Priority Email (4h) | Shared Slack (2h) | Dedicated TAM & 15m Critical SLA |
Got Questions? We Have Answers
Everything you need to know about billing, BYOK model keys, token limits, and infrastructure deployment.
Ready to Orchestrate at Peak Performance?
Start orchestrating with your existing API keys in under 5 minutes. No vendor lock-in, zero latency compromise.