
Key Features
Unified API for 1000+ LLMs
Multimodal & Audio APIs
Native SDK Compatibility
Load Balancing & Fallbacks
Semantic Caching
Batch APIs
Access Control & API Keys
Rate Limiting
Budgets & Cost Tracking
Guardrails
Observability & Logs
Prompt Management
MCP Registry
Centralized MCP Auth
Virtual MCP Servers
Agent Registry
Skills Registry
SKILL.md instructions for agents and IDEs.Flexible Deployment
Supported Model Providers
We integrate with 1000+ LLMs through the following providers.

















Supported APIs
The following accordions summarize provider support for each gateway endpoint. Each section links to the full guide for that API (same order as Supported APIs in the sidebar).- ✅ Supported by Provider and Truefoundry
- Provided by provider, but not by Truefoundry
- Provider does not support this feature
Chat Completion (/chat/completions)
Chat Completion (/chat/completions)
Embedding (/embeddings)
Embedding (/embeddings)
Batch (/batches)
Batch (/batches)
Fine Tune
Fine Tune
Model Response (/responses)
Model Response (/responses)
Image Generation (/images/generations)
Image Generation (/images/generations)
Image Edit (/images/edits)
Image Edit (/images/edits)
Image Variation (/images/variations)
Image Variation (/images/variations)
Text To Speech
Text To Speech
Audio Translation
Audio Translation
Speech to Text
Speech to Text
Live / Realtime API
Live / Realtime API
Files (/files)
Files (/files)
Rerank (/rerank)
Rerank (/rerank)
Moderation (/moderations)
Moderation (/moderations)
Compaction API
Compaction API
Messages API
Messages API
Proxy API (/proxy)
Proxy API (/proxy)
Deployment Options
You can run the AI Gateway as fully managed SaaS, keep LLM request–response data in your own object storage while Truefoundry operates the gateway, or host the gateway plane (and optionally more of the stack) in your cloud or on-prem for stricter data residency and control. Each option differs in who hosts infrastructure, where traffic flows, and pricing tier. Read the full comparison—including a scenario table, diagrams, and operational notes—in AI Gateway deployment options. For background on how the gateway fits the platform, see gateway plane architecture. To start on managed SaaS, follow the quick start.Frequently Asked Questions
What's the performance impact of using the gateway?
What's the performance impact of using the gateway?
is closer to your users.

AI Gateway on the edge, close to your applications for optimal performance
Can I deploy the gateway on-premise?
Can I deploy the gateway on-premise?
How do I integrate my self-hosted models?
How do I integrate my self-hosted models?
Can I use the gateway without the full MLOps platform?
Can I use the gateway without the full MLOps platform?