zintus
How it worksFeaturesSupported AIsFAQInstallDownloadPricingDocs
StarOpen app
One subscription. Exact tokens. Zero surprises.

The AI router that tells you exactly what you bought.

Most AI subscriptions hide their limits behind rolling windows, compute credits, and message caps. Zintus gives you a number. It doesn't change.

Start routing freeSee pricing

No account required · Keys stay on your device · Token balance always visible

Free forever with your own API keys · Managed keys from $15/mo →

zintus — terminal
 
 
 
 
build from source · npm publish coming soon:
CerebrasGroqGoogle AIDeepSeekMistralTogetherFireworksOpenRouterxAICohereOllamaLM StudioCerebrasGroqGoogle AIDeepSeekMistralTogetherFireworksOpenRouterxAICohereOllamaLM Studio

ROUTING LOGIC

  • Cerebras
  • Groq
  • Gemini

QUOTA USAGE

80% used · 1.2M tokens remaining

zintus.ai/chat — 128ms latencyACTIVE
Explain how neural networks learn
Type a command or query…⌘↵
0+ providersDirect + meta-router catalogDIRECT + BYOK CATALOG
0+ modelsCurated and growing weeklyREFRESHED WEEKLY
< 0ms routingIn-process quota checkIN-PROCESS QUOTA CHECK
0% markupOn your own API keysON YOUR OWN API KEYS

How it works

Smart routing in three steps.

01

Add your keys

Drop in your API keys. They're stored in your OS keychain — never sent to our servers. Add one provider or twelve. Each one multiplies your free daily quota.

$ zintus keys add
$ zintus keys list
# stored in your OS keychain
02

The router decides

Zintus checks quota across all providers in real time. Short prompt? Routes to the fastest model. Long context? Picks the one with a 2M token window. Code heavy? Goes to the best coder at T0 price.

$ zintus chat "summarize this contract"
# → widest-context model · 2M tokens
03

See exactly what it cost

Every response shows the model used, tokens consumed, and dollar cost. Your monthly balance updates in real time. No surprises at the end of the month.

model: fast model · tokens: 1,247
cost: $0.004 · balance updated
# shown after every response

Features

A router built around what you actually bought.

Routing, billing, and key handling are built around one idea: you always know which model answered, what it cost, and how much budget is left.

Smart routing, real savings

Every request is classified and sent to the best-value provider — short queries to the fastest model, long contexts to the widest window, reasoning to the strongest reasoner.

$classify & route request
gemini-2.0-flashlong context240ms
deepseek-v3code180ms
llama-3.3-70breasoning95ms

Exact token balance

Your monthly budget is a fixed number, shown before you subscribe and after every message. It never shrinks at peak hours or changes with the model you use.

Starter · 1M tokens847K left
Fast models−1 / 1K
Advanced models−5 / 1K

No silent downgrades

Zintus shows which model handled every request. If your budget runs low we warn you — we never quietly swap what you're getting.

Your keys stay on your device

BYOK keys live in your OS keychain or the browser's encrypted storage and go straight to the provider. Zintus never sees them.

90+ providers, one interface

20+ direct integrations plus 70+ more via a model-aggregator key. One config, one dashboard, one token balance.

BYOK frontier on any plan

Add your own key for any frontier model on any paid tier — direct provider rates, zero markup, off your own credits.

Transparency

Every route. Every token. Receipted.

Routing logIllustrative sample
TimeModel routeProviderTokensLatencyCost
09:42:01gemini-2.0-flashEdge optimizedGoogle1,203240ms$0.0004
09:41:56deepseek-v3Fallback triggeredDeepSeek2,847180ms$0.0009
09:41:48llama-3.3-70bCheap route enabledCerebras64095ms$0.0001
09:41:33claude-haikuDirect ingressAnthropic892310ms$0.0007
09:41:12qwen-2.5-72bCheap route enabledGroq1,54088ms$0.0003
gemini-2.0-flash
Edge optimized
Google
Time
09:42:01
Tokens
1,203
Latency
240ms
Cost
$0.0004
deepseek-v3
Fallback triggered
DeepSeek
Time
09:41:56
Tokens
2,847
Latency
180ms
Cost
$0.0009
llama-3.3-70b
Cheap route enabled
Cerebras
Time
09:41:48
Tokens
640
Latency
95ms
Cost
$0.0001
claude-haiku
Direct ingress
Anthropic
Time
09:41:33
Tokens
892
Latency
310ms
Cost
$0.0007
qwen-2.5-72b
Cheap route enabled
Groq
Time
09:41:12
Tokens
1,540
Latency
88ms
Cost
$0.0003

Provider spend share

Cerebras42%
Groq27%
Google19%
DeepSeek12%

Requests by latency

<50
50–100
100–150
150–200
200–300
300–500
500+

Take it everywhere.

Zintus works on your terminal, browser, desktop, and phone. One config, all devices.

CLIbuild from source (Bun)zintus chat 'hello'Webzintus.ai/chatWorks in any browserMobileiOS & AndroidControl your gateway remotely

Desktop app available for macOS, Windows, and Linux (beta). See all downloads →

Supported providers

20+ providers. 90+ models. Smarter every month.

Free tier routes across 15+ providers with real-time quota tracking. Add a model-aggregator key to reach 400+ models instantly. Managed tiers use Zintus-provided keys across our curated routing roster.

Ultra-fast inference1M tok/day
70B · low latency1K req/day
8B · high volume14.4K req/day
Fast multimodal1.5K req/day
Open reasoningMIT free
Model aggregator300+ models
Ultra-fast inference1M tok/day
70B · low latency1K req/day
8B · high volume14.4K req/day
Fast multimodal1.5K req/day
Open reasoningMIT free
Model aggregator300+ models
Enterprise NLP1K/month
Open weights~1B tok/mo
Serverless inference$1 free
Frontier trialtrial
Local runtimeunlimited local
Local desktopunlimited local
Enterprise NLP1K/month
Open weights~1B tok/mo
Serverless inference$1 free
Frontier trialtrial
Local runtimeunlimited local
Local desktopunlimited local
Tier 0 — Default94–99% margin · < $0.30/M
Tier 1 — Capable70–85% margin · $1–3/M
Tier 2 — CeilingTop-tier reasoning model · hard cap
Browse the full catalog — every provider and model →

For developers

Up in 60 seconds

Build the CLI from source with Bun (npm publish coming soon). Keys are stored in your OS keychain — never sent anywhere.

TERMINAL
$ bun install
$ bun run --filter zintus build
$ zintus chat

Questions & answers

Frequently asked questions

Is it really free?

Yes. Zintus uses the free tiers of each AI service, and the app itself is free to use. You only pay if you choose to upgrade a provider's plan yourself.

Are my API keys safe?

Yes. Your keys are stored securely on your own device — in your browser's encrypted storage, your computer's keychain, or your phone's secure store. They are never sent to our servers.

Do I need to be a developer?

No. The web and desktop apps work like any normal chat app. Developers also get a command-line tool and a self-hostable gateway if they want them.

What happens when a free AI runs out?

Zintus tracks each service's usage and automatically switches to another available one when a provider is exhausted.

Which AIs are supported?

20+ providers today across direct integrations — ultra-fast inference, large reasoning models, fast multimodal, and enterprise NLP in the cloud — plus local runtimes on your own computer. Add a model-aggregator key to reach 300+ more models.

Is it open source?

Zintus is source-available under the Business Source License 1.1 — you can self-host it for personal or internal use at no cost. It's not an OSI open-source license; commercial hosting or resale needs a separate license.

How is this different from a paid API proxy?

A typical API proxy adds a markup on every call and sits between you and the provider. Zintus is a local router — your requests go directly to each provider from your machine, with zero markup and no middleman. We add smart quota tracking, automatic fallback, and token compression on top.

Can I use it with a single API key?

Yes. Even one key gives you up to 1,500 free requests a day with search grounding built in. Add more keys over time — each one multiplies your free quota.

Is my data private?

Your prompts go directly from your device to the AI provider — Zintus never sees them. API keys are stored in your OS keychain or browser's encrypted storage, not on our servers.

Does it work offline?

Yes. When all cloud providers are exhausted, Zintus automatically falls back to a local runtime on your own machine. You always get a response.

Does it support tool calling, JSON output, and images?

Yes. Define tools once and Zintus maps them to each provider's native function-calling format, routing only to models that support it. Ask for structured output and you get JSON — schema-constrained where the provider guarantees it, best-effort JSON mode elsewhere. You can also send images to vision-capable models; EXIF and GPS metadata are stripped on your device first.

How do I know how much I'm saving?

Zintus keeps a usage ledger and values every free-tier token against what a metered API would have charged for the same model. The dashboard shows the running estimate — it's an approximation for display, not a bill, but it's based on your real usage.

What's the difference between BYOK and managed tiers?

On the free BYOK tier, you supply your own API keys. They stay on your device and requests go directly to providers — Zintus never sees them. On managed tiers ($15–$199/mo), Zintus provides the keys. You get an exact monthly token budget and we handle routing, quota tracking, and automatic fallback.

Can I use frontier models on a managed plan?

Yes — via BYOK on top. Add your own provider key to any paid tier and we route frontier requests through your key at zero markup. Your managed budget handles everything else. You get both: a predictable monthly budget and access to the most powerful models available.

Why do you charge in tokens instead of messages or credits?

Tokens are the actual unit of work in every AI system. Messages and credits are abstractions layered on top — designed to make limits harder to compare. We skip the abstraction. You see the real number.

What happens when I run out of tokens?

You'll see your balance in real time throughout the month. When it runs low we warn you. When it hits zero, requests stop — no surprise overages. You can top up with an add-on, or add your own API key to continue at direct provider rates with no Zintus markup.

Start today

Built for developers who hate paying for AI.

All the power. None of the bills. Your keys, your data, your machine.

Start routing freeSee pricing
zintus

Source-available AI router. BYOK. Zero markup. Routes intelligently across 12 free providers.

Product

PricingChangelogChatDesktop App

Developers

DocsDevelopersGitHubCLI InstallAPI Referencellms.txt

Company

AboutBlogContactTwitter/X

Legal

PrivacySecurityLicense (BUSL-1.1)Terms
© 2026 Zintus · YS Ventures LLC · Business Source License 1.1Made with ❤️ in Pittsburgh, PA