Ship AI at full capability.
One unified inference API across every model, hosted or local, with the registry that keeps each one configured, priced, and routable. Plug in once and build.
Use Switchboard
Switchboard can be integrated with your solo harness setup, a team on one balance, or an app delivering agentic software at scale to millions of users.
The platform
One bill. Every model.
A unified inference API across every provider: send the format you already use, route by model id, settle it all on one invoice. No provider keys, no juggling accounts.
Explore Switchboard SwitchboardLocalThe same API, on-device.
Run open models locally through the exact Switchboard interface. Zero metering, no data leaving the machine, and one line to fall back to hosted models when you need them.
Explore SwitchboardLocal LinecardThe registry behind routing.
Pricing, capabilities, and provenance for every model, human-verified before it goes live. Linecard is what Switchboard reads to route, so every model runs at full capability.
Browse LinecardManaged inference solutions for solo engineers, teams, and companies building AI applications.
Here is what each capability does, and who it works hardest for.
Feature
Description
Target use
Routing
One request shape reaches every live model in the catalog; change models by changing the model id. Switchboard speaks each provider's native format, so nothing is lost in translation: streaming, tool calling, thinking, and caching all work at the model's full capability.
Your agent, your backend, and your app all send the format they already speak.
Metering
Every token measured as it streams: input, output, cache writes, cache reads, and reasoning, itemized at provider-exact rates with cached-prompt discounts included.
You know what any request cost the moment it finishes.
Control
Spend caps, rate limits, and model policy set per user and enforced at the router, not in your code.
A budget for every teammate, a ceiling for every end user.
Research
Capabilities, configuration, and per-token prices for every model, human-verified in Linecard before they go live. Pick by what a model can do, not by what its name suggests.
Compare the new frontier model against your current one before routing a single request to it.
Analytics
Usage by model, by user, and over time, with typed error records in each user timeline.
You see who used what, what it cost, and when a request failed, why.
Key management
Switchboard mints and manages keys for your application and every end user. Nothing ships in a binary, nothing gets pasted, and rotation happens over the wire.
No key sprawl across the team, no credentials inside the app you ship.
One bill
One prepaid balance covers every provider, every debit lands on a per-user ledger, and it all settles to a single itemized invoice.
Whether it is your own usage or a thousand users, the bill is one document that adds up.
The same API, on-device.
Run open models locally through the exact Switchboard interface: same request shape, same client, and nothing leaves the machine. When a task outgrows the local model, fall back to hosted models with one line.
SwitchboardLocal ships in the Swift SDK on macOS. The pi extension routes hosted models only.
Explore SwitchboardLocal-
Zero metering
Local tokens are free tokens. Nothing books to your balance.
-
Private by construction
Prompts and outputs never leave the machine.
-
One-line hosted fallback
Same client, so escalating to a frontier model is a model id change.
Nobody should hand-maintain a model catalog.
Maintaining it yourself
Every provider is different
Request quirks and tool-calling rules vary by provider. What works against one model quietly fails against the next.
Capabilities are a moving target
Vision, tools, caching, context limits: every model supports a different set, documented in a different place.
Pricing drifts without notice
Per-token rates change under you, and the first place you find out is the bill.
New models mean new work
Every launch is another round of setup and testing before anyone on your side can use it.
With Switchboard
Instant models, on demand.
A new model shows up on your account the day it ships, already configured and priced. Pick it by what it can do, change the model id, and keep working. There is nothing to update and nothing to maintain.
- New models arrive ready to use
- Every model runs at full capability
- Prices are always current, on every invoice