← sovetrust gateway · Soverouter · in sovetrust
Local-first AI · 80 / 20

Your AI runs on your machine. The cloud is the exception.

Soverouter keeps roughly 80% of inference local — private, zero-cost — and routes only the hard 20% to the cheapest capable cloud model. The inverse of cloud-first assistants.

Two ways to run

Pick your surface

With agents

Soverouter Agents

Agentic local AI — tool-calling, memory, and automation, powered by local agent models with cloud fallback for hard reasoning.

  • Hermes 3 and other local agents (tool use, planning)
  • Persistent memory, sandboxed actions
  • Credentials never leave the device (read-blocked)
  • Cloud fallback only when local can't reason deep enough
agents: Hermes 3 · Qwen2.5-Agent · Llama-Tool
Without agents

Soverouter Core

Just fast, cheap AI. No agent overhead — one endpoint that auto-routes each request to the lowest cost/latency model that can do the job.

  • 3-tier auto-routing (local → free cloud → paid)
  • Cost/latency arbitrage on every request
  • OpenAI-compatible endpoint — drop-in
  • Bring your own keys or use platform credits
models: Llama · Qwen · Groq · GPT-4o · Claude
Live routing

See where a request would go

Set the inputs and compute a route plan.
Pricing · undercut the cloud-first crowd

Plans

Rivals charge ~$20/mo for cloud-first. Local-first means your free tier actually costs us cents — so we can price low. Billing by Stripe.

Developers · SDK

Drop-in, OpenAI-compatible

Point the Sove SDK at your routing endpoint. It plans the tier, runs local when possible, and falls back automatically.

Account

Your plan & credits

Sign in to view your plan, Sovereign Credits, and routing history.