Logo SVG copied to clipboard
Offer 🎁 Talk to us and get $25 of Gemini credit

Take back the control of your LLM calls

Manifest is an open source LLM gateway available on the cloud or on your own infrastructure.

  • Budgets and rate limits
  • Spend per engineer
  • Bring your own providers
  • Local models

Trusted by engineers who ship at

Manifest is an AI model gateway that you can trust

Model fallbacks

Automatically retry on another provider when one goes down.

Full body logs

Inspect the complete request and response body for every call.

Autofix

Broken requests get repaired on the fly, before your agent sees the error.

Local models

Route to Ollama, LM Studio, llama.cpp or any local server. Decide what goes to the cloud and what stays on your machine.

Bring your own key

Bring your own keys for every provider and pay usage directly. Manifest allows API keys from all popular providers and OAuth tokens for subscriptions.

Easy parameter setup

Set the right model and parameters per route without touching the code. Set AI model parameters visually.

Cost visualization

See every dollar spent, broken down by agent, key and provider.

SSO / SAML

Bring your own identity provider and manage team access through SSO and SAML.

Custom setup

Need something specific? We tailor the deployment and configuration to your team.

Don't get locked into a single provider

Simultaneously use different kinds of model or inference providers based on the query. Manifest does not limit you to a restricted list of providers.

OpenAI Anthropic Google Mistral Meta DeepSeek xAI Groq Cohere NVIDIA Cerebras Hugging Face Ollama GitHub Alibaba Cloud MiniMax SiliconFlow Cloudflare Workers AI Arcee AI

API key providers

Bring your own key from all the main providers to get instant access to all their models. Pay by the usage directly to the provider.

Subscription providers

Already paying for a monthly subscription? Connect it to Manifest to use those quotas first, fallback to pay-as-you-go only when limits are exceeded.

Custom providers

Plug in any OpenAI-compatible or Anthropic-compatible provider that exists out there! We don't limit you to the providers that we know only.

Local models

Run open-weight models at home! Manifest handles Ollama, LM Studio and llama.cpp as first-class providers so you can run them on your own infrastructure.

One endpoint for the whole team. Start free.

Frequently asked questions

How do shared budgets work?

Set a cap per key, daily, weekly or monthly. When a key reaches it, Manifest stops routing on that key and sends an alert. Nobody discovers a $3,000 weekend on next month's invoice.

Can I see who spent what?

Give every engineer their own key. The dashboard breaks cost down by agent, key and provider, so you can see which model is eating the budget and who is calling it.

Do we need Enterprise to run a team?

Free and Pro are built for one person. Seats, SSO/SAML, audit logs and custom retention come with Enterprise, on our cloud or on your own servers.

Do you offer an SLA?

Enterprise includes a contractual uptime SLA, with on-premise, VPC or private cloud install. Free and Pro run on best effort.

How do we onboard the team?

Change one base URL. Manifest speaks both the OpenAI and Anthropic APIs, so Claude Code, OpenClaw, the OpenAI SDK and the Vercel AI SDK need no code change. Hand each engineer a key and they keep the setup they already have.

Can we keep our own provider keys?

Yes. Bring a key per provider and pay each provider directly. We take no cut of your tokens. Subscriptions work too (Claude, ChatGPT, GitHub Copilot and others), and the self-hosted version routes to Ollama, LM Studio or llama.cpp when a prompt should not leave your network.