[01] Products
Pick your control plane
One platform, two deployment models. Both serve the same OpenAI-compatible API from your own hardware — the difference is who runs the control plane.
[02]Deployment models
[03]Shared stack
01
Conduit
Self-hostable agent connecting vLLM, llama.cpp and other engine servers on your machines.
02
Endpoints
Public or private inferencing endpoints with composable routing to your sources.
03
Models
Serve any local model; curated recommended-models catalog included.
04
API
OpenAI-compatible REST surface — drop-in for existing agents and tooling.
05
Tool services
Server-side tool interception and MCP tool services wired into requests.
06
Console
Sources, models, endpoints and spend managed from one control panel.