meta-llama/Llama-3.3-70B-Instruct, deepseek-ai/DeepSeek-V3-0324, and openai/gpt-oss-120b. Use Custom Endpoints in the AI Gateway to register Crusoe once, keep the Crusoe API key on the AI Gateway, and let your applications call models through a single TrueFoundry API key with access control and tracing.
Prerequisites
1
Create a Crusoe Cloud account
Sign up at console.crusoecloud.com if you do not already have an account.
2
Generate a Crusoe API key
In the Crusoe Cloud Console, go to Admin → Security → Intelligence API keys and click Create.Store the key securely. You will paste it into the AI Gateway’s Custom Headers configuration — not in client application code.
3
Get your TrueFoundry API key and gateway URL
You need a TrueFoundry API key (
TFY_API_KEY) and your gateway base URL to call models through the AI Gateway. See Gateway base URL and Authentication.Adding models
Add Crusoe to the AI Gateway using Custom Endpoints.1
Create a Custom Endpoint model account
In the TrueFoundry dashboard, go to AI Gateway → Models → Custom Endpoints and click Add Custom Endpoint.In Configure Account, set:
- Name (Account Name):
crusoe— this is your model account name and appears as the first path segment in every request URL ({providerAccountName}) - Endpoint Type:
None - Header Auth: keep disabled
2
Add a Crusoe endpoint
On the Endpoints step, add an integration and configure:
- Display Name:
crusoe_managed_inference— this is your custom endpoint display name and appears as the second path segment in every request URL ({endpointName}) - Base URL:
https://api.inference.crusoecloud.com(no trailing slash)
Authorization:Bearer <CRUSOE_API_KEY>Content-Type:application/json
The Authorization header here is sent from the AI Gateway to Crusoe. Your application should only send the TrueFoundry API key to the AI Gateway.
3
Set access control and save
On the Access Control step, choose who can manage and use this model account, then Save.
- Manager: users/teams who can edit or delete the custom endpoint
- User: users/teams who can call the endpoint (for example,
everyone)
Inference
Once saved, call Crusoe through the AI Gateway’s proxy-api path. URL shape and path rules are documented under Custom Endpoints.How the request URL is built
The AI Gateway URL has two values you configure in the dashboard — they are not Crusoe model IDs:
Full URL pattern:
The AI Gateway forwards the request to:
Choosing a Crusoe model
Crusoe hosts open-weight models such asmeta-llama/Llama-3.3-70B-Instruct, deepseek-ai/DeepSeek-V3-0324, and openai/gpt-oss-120b. You do not need a separate custom endpoint per model — set the model field in the JSON request body to the Crusoe model ID you want. See the Crusoe serverless inference docs for the full list.
Replace
crusoe and crusoe_managed_inference in the URL with your own Account Name and Display Name if you used different values during setup.Supported APIs
Chat Completions
Chat Completions
Before you start: Replace
{GATEWAY_BASE_URL} with your gateway base URL (how to find it) and set TFY_API_KEY to your TrueFoundry API key.Request headers (client → gateway)
OpenAI Python SDK
Because Crusoe is OpenAI-compatible, you can point the OpenAI SDK at the AI Gateway proxy path:cURL
Support scope: Custom Endpoints proxy requests transparently to Crusoe. See the Crusoe Managed Inference docs for supported models and request fields. For gateway limitations on Custom Endpoints (HTTPS, streaming, and so on), see Custom Endpoints.