FIM Gate Documentation
FIM Gate is an LLM API aggregation gateway — call models from multiple providers with a single API key, compatible with the OpenAI, Claude, and Gemini APIs.
FIM Gate is an LLM API aggregation and forwarding service: with a single API key and a single endpoint, you can call models from OpenAI, Claude, Gemini, and other providers. For most existing projects, all you need to do is change the base URL to https://api.gate.fim.ai — the rest of your code stays the same.
Core capabilities
- OpenAI API compatibility: standard endpoints such as
/v1/chat/completionswork out of the box, so any SDK, framework, or client that supports a custom base URL can connect directly. - Native multi-provider protocols: beyond the OpenAI format, FIM Gate also forwards native Claude Messages API and Gemini GenerateContent API requests, so you can keep using each provider's official SDK.
- Unified key management: create and disable API keys and review usage details in the Console — all models share a single key system.
- Pure forwarding architecture: the service only forwards requests and meters usage; it does not store conversation content.
- Streaming output: all endpoints support SSE streaming responses, ideal for building real-time chat experiences.
Get started in three steps
- Sign up and create an API key in the Console.
- Change your request URL to
https://api.gate.fim.ai/v1(OpenAI-compatible format). - Include
Authorization: Bearer $FIM_API_KEYin the request headers and make calls as usual.
curl https://api.gate.fim.ai/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $FIM_API_KEY" \
-d '{
"model": "gpt-4.1",
"messages": [{"role": "user", "content": "Introduce yourself in one sentence"}]
}'Documentation guide
| What you want to do | Where to look |
|---|---|
| Make your first call and get a request working | Quickstart |
| Figure out which endpoint to use | Base URL |
| Understand how costs are calculated | Billing |
| Create API keys and configure third-party clients | Keys & Tokens |
| Compatibility between the chat and responses endpoints | Chat API |
| Model name suffixes, Gemini web search, and other tips | Special Usage |
| Control the thinking process of Claude / Gemini | Reasoning Models |