v1.0 — Public Launch
nRouter v1.0 is live — managed LLM gateway with the full Google Vertex AI lineup, full dashboard, credit-based billing, AI guardrails, prompt management, A/B testing, team-scoped budgets, and observability — every feature on every tier.
nRouter v1.0 is publicly available. This is the first generally-available release of our managed LLM gateway: one API key, one bill, every feature unlocked on every tier — with multi-tenancy, billing, and a governance layer built in.
What ships at v1.0
Models
The full Google Vertex AI lineup is live on day one — Gemini chat models, embeddings, Imagen image generation, and Veo video. OpenAI, Anthropic, AWS Bedrock, Azure OpenAI, Cohere, Mistral, and Meta integrations shipping in the weeks after launch — every new model lands in days, not quarters, because we extend the open-source core rather than fork it. The live count is always shown on the models page and is derived from the router, never hardcoded.
Smart routing
Pick a routing strategy (usage, latency, cost, simple-shuffle, least-busy). Configure fallback chains that transparently retry on backup models when a provider fails. Per-org retry counts, timeouts, and cooldowns. Tag-based filtering for capability-aware routing (vision, code, long-context). Every routing decision captured in observability.
AI guardrails
Five guardrails on every request, included on every plan from day one — no Enterprise paywall:
- PII redaction (Microsoft Presidio)
- Prompt-injection detection on adversarial corpora
- API-key + secret scanning on prompts and completions
- Abuse blocking
- Response scanning
Configurable scope: organization > team > virtual key, with override semantics.
Prompt management
Server-side prompt templates with versioning, A/B testing, Jinja2 variables, and per-template cost tracking. Run experiments deterministically with traffic-split variants — same hash, same variant, every time.
Team management & budgets
Multi-org support with role-based access (Owner, Admin, Member, Viewer). Per-team and per-virtual-key spend visibility. Hard and soft budget limits at any scope. Invitation flow with single-org-per-user enforcement.
Credits as the product
Buy credits from $1. Tier 1 (PAYG, 4% platform fee) — $10 free credits on signup. Tier 2 ($100/mo, 2% fee). Tier 3 ($1,200/yr, 0% fee). Enterprise (0% fee, custom). Every dollar you commit goes to credits — platform fee is added on top, never deducted. Reserve+settle pattern under Postgres advisory locks means denied requests cost zero credits.
Pricing update (September 2026): the launch plans above, and the July 2026 Pro plan with a 0% fee, are no longer sold. Every self-serve plan (Pay as you go, Starter, Pro, Max) now pays 4% of the credits, charged on top, and there are no free signup credits. See nrouter.ai/pricing.
Credits update (July 2026): the signup credit grant described above has been retired. There is no free tier — every account starts with a paid checkout: a $5 minimum credit purchase on Pay as you go (platform fee charged on top), or a Pro subscription. New customers receive $10 bonus credits on their first purchase, so a first $5 load becomes $15 in credits.
Plans update (September 2026): the bonus on a first purchase described in the July note is no longer offered to new signups. Current plans and pricing: nrouter.ai/pricing.
Observability
Request logs with model/status/time filtering and expandable detail rows. Logging callbacks to Langfuse, Datadog, S3, and Slack. Configurable per-org data policy: zero-logging, metadata-only, full-logging, or PII-redacted. 90-day request log retention; longer on Enterprise.
Security architecture (built-in, not gated)
Postgres Row-Level Security on every nRouter table — tenant isolation enforced at the database, not the application. Customer LLM traffic uses only virtual keys (sk-nrouter-…), never master keys. Hashed virtual-key storage; plaintext shown once. Encryption: TLS 1.2+ in transit (HSTS preloaded), AES-256 at rest. Immutable audit trail on every administrative action. Reserve+settle credit ledger means negative balances are blocked at the database layer.
Compliance posture
SOC 2-aligned controls (active). ISO 27001-aligned controls (active). GDPR-compliant. HIPAA-eligible (BAA available on Enterprise). PCI delegated to Stripe. Underlying infrastructure (Google Cloud Run, Supabase) is SOC 2 Type II certified.
Get started
pip install nrouter-sdk # or use the OpenAI SDK with our base URLSign up at nrouter.ai — Pay as you go starts with a $5 minimum credit purchase, the platform fee on top. Read the API spec at /docs. Talk to us at /contact.
Platform release verification & compatibility standards
All nRouter release notes reflect changes deployed across our global edge routing infrastructure. Below is our baseline deployment architecture and verification criteria.
Zero-Downtime Rolling Delivery
Releases deploy via immutable container digests with automated health verification. Existing streaming responses complete uninterrupted while new requests seamlessly route to updated instances.
Drop-In API Backward Compatibility
Our interface remains strictly compatible with standard OpenAI client SDKs and wire formats. Enhancements degrade gracefully or fail open with explicit error reporting to protect calling workloads.
How can engineering teams verify capabilities in this release?
Teams can inspect currently supported capabilities and active model deployments programmatically via the authenticated /capabilities endpoint, or review live latency profiles at /models.
Where can edge cases or regressions be reported?
If your team experiences unexpected response behavior, report the issue with the relevant x-nr-request-id header to support@nrouter.ai or via the community Slack at /community.