Valni

Ship AI at full capability.

One unified inference API across every model, hosted or local, with the registry that keeps each one configured, priced, and routable. Plug in once and build.

Switchboard

Managed inference solutions for solo engineers, teams, and companies building AI applications.

Here is what each capability does, and who it works hardest for.

Routing

One request shape reaches every live model in the catalog; change models by changing the model id. Switchboard speaks each provider's native format, so nothing is lost in translation: streaming, tool calling, thinking, and caching all work at the model's full capability.

Your agent, your backend, and your app all send the format they already speak.

PersonalTeamProduct

Metering

Every token measured as it streams: input, output, cache writes, cache reads, and reasoning, itemized at provider-exact rates with cached-prompt discounts included.

You know what any request cost the moment it finishes.

PersonalTeamProduct

Control

Spend caps, rate limits, and model policy set per user and enforced at the router, not in your code.

A budget for every teammate, a ceiling for every end user.

TeamProduct

Research

Capabilities, configuration, and per-token prices for every model, human-verified in Linecard before they go live. Pick by what a model can do, not by what its name suggests.

Compare the new frontier model against your current one before routing a single request to it.

PersonalTeamProduct

Analytics

Usage by model, by user, and over time, with typed error records in each user timeline.

You see who used what, what it cost, and when a request failed, why.

TeamProduct

Key management

Switchboard mints and manages keys for your application and every end user. Nothing ships in a binary, nothing gets pasted, and rotation happens over the wire.

No key sprawl across the team, no credentials inside the app you ship.

TeamProduct

One bill

One prepaid balance covers every provider, every debit lands on a per-user ledger, and it all settles to a single itemized invoice.

Whether it is your own usage or a thousand users, the bill is one document that adds up.

PersonalTeamProduct
Get started with Switchboard
SwitchboardLocal

The same API, on-device.

Run open models locally through the exact Switchboard interface: same request shape, same client, and nothing leaves the machine. When a task outgrows the local model, fall back to hosted models with one line.

SwitchboardLocal ships in the Swift SDK on macOS. The pi extension routes hosted models only.

Explore SwitchboardLocal
  • Zero metering

    Local tokens are free tokens. Nothing books to your balance.

  • Private by construction

    Prompts and outputs never leave the machine.

  • One-line hosted fallback

    Same client, so escalating to a frontier model is a model id change.

Linecard

Nobody should hand-maintain a model catalog.

Maintaining it yourself

Every provider is different

Request quirks and tool-calling rules vary by provider. What works against one model quietly fails against the next.

Capabilities are a moving target

Vision, tools, caching, context limits: every model supports a different set, documented in a different place.

Pricing drifts without notice

Per-token rates change under you, and the first place you find out is the bill.

New models mean new work

Every launch is another round of setup and testing before anyone on your side can use it.

With Switchboard

Instant models, on demand.

A new model shows up on your account the day it ships, already configured and priced. Pick it by what it can do, change the model id, and keep working. There is nothing to update and nothing to maintain.

  • New models arrive ready to use
  • Every model runs at full capability
  • Prices are always current, on every invoice
Browse Linecard