Conifer — run AI locally, route the rest
conifer
Conifer routes to the Pareto frontier
Works with the cloud, self-hosted models, and your own hardware.
Install the runtime
No charge. The runtime, the router and local inference are free.
One key to the gateway
Over 200 models behind
api.conifer.build: the frontier labs and the open-weights field, Qwen, DeepSeek, GLM, Kimi, Llama and the rest. Or bring your own provider keys.No upcharge
Cloud tokens bill at the model’s own rate and nothing on top. Your own keys and your own hardware, through any llama-compatible endpoint or our engine, cost nothing.
Dynamic query-based routing.
Works anywhere.
- Claude Desktop
- Codex Desktop
- Antigravity
- Terminal
- cmux
- VS Code
- OpenCode
- Pi
- Zed
- Cursor
The harness, the model, the provider and the cache, handled for you.
Each query gets those four chosen for it as it arrives. All you send is the text.
› refactor the auth middleware, keep the tests green
askanswer
harnessmodelprovidercache strategy
Claude CodeKimi K3Togethercache read









