You can proxy requests to Cerebras AI models through AI Gateway by creating AI Model Provider and AI Model entities. This reference documents all supported AI capabilities, configuration requirements, and provider-specific details needed for proper integration.
Cerebras provider
Upstream paths
AI Gateway automatically routes requests to the appropriate Cerebras API endpoints. The following table shows the upstream paths used for each capability.
|
Capability |
Path template |
Description |
Upstream path or API |
|---|---|---|---|
| Generate |
/chat/completions, /completions, or /responses
|
Text generation for chat completions and responses |
/v1/chat/completions
|
Supported capabilities
The following tables show the AI capabilities supported by the Cerebras provider when configuring AI Models.
By default, AI Gateway uses the path templates shown in the tables below (e.g.,
/chat/completions,/embeddings, etc.). To customize these paths, configure theconfig.pathsfield in your AI Model entity. Custom paths take the form{configured_path}/{template_path}— for example, if you set a custom path of/v2, requests to/embeddingswould be routed to/v2/embeddings.
Text generation
Support for Cerebras text generation capabilities:
|
Capability |
Streaming |
Model example |
Path template |
Min version |
|---|---|---|---|---|
| generate | Supported | llama-3.3-70b |
/chat/completions, /completions, or /responses
|
2.0 |
Cerebras base URL
The base URL is https://api.cerebras.ai/{capability_path}. The {capability_path} is determined by the AI capability.
AI Gateway uses this URL automatically. You only need to configure a URL if you’re using a self-hosted or Cerebras-compatible endpoint, in which case set the upstream_url option in your AI Model configuration.
Configure Cerebras
To use Cerebras with AI Gateway, configure a new AI Model Provider. You can then access supported AI Models from Cerebras.
Here’s a minimal configuration for chat completions: