Skip to main content

Enabling new Models

To see available Models and enable them for use, navigate to the Models page in AI Gateway.
Models page with the Model Filters sidebar listing Location, Access, Providers, Status, Features, Context length, Intelligence index, Input price / 1M tokens, and Owner, the modality tabs including Classify, and columns for Name, Status, Model Access, Price, Intelligence, Capabilities, and Released.

The Models page showing the filter sidebar, modality tabs, and model columns.

Each model row carries the following columns. Status and Model Access are visible to workspace admins only. Use Sort: Newest to reorder by Newest, Pricing: Low to High, Pricing: High to Low, Context: Low to High, Context: High to Low, Intelligence: High to Low, Intelligence: Low to High, or Max Output Tokens. Use Columns to show or hide individual columns.
Use the Status toggle to Enable a model for use with the AI Gateway.

Row actions

Hover a model’s Name cell to reveal that row’s actions: Private models carry a Delete action in the row menu. Models onboarded through providers that support editing, such as OpenAI Compatible, Azure, and Google, also show Edit, which reopens the provider wizard.

Filters

Use the modality tabs at the top of the list to scope models by type: All Text Image Audio Speech Embedding Moderation Rerank Classify The sidebar provides additional filters: To enable a model, toggle it on. It will immediately be available to call with the AI Gateway.

Enable or Disable Models via the API

Models can also be enabled and disabled programmatically using the Models API. This is useful for CI/CD pipelines, automation scripts, or infrastructure-as-code workflows. These endpoints sit on the management plane and authenticate with a Management Key that has workspace-model write access. A standard API Key cannot be granted the workspace-model domain and is rejected with 403. Enable a model:
Disable a model:
Both endpoints return 204 on success. Re-enabling an already-enabled model or disabling an already-disabled model is idempotent and also returns 204. When Enforce enabled models is turned on in General Settings, only models enabled through the dashboard or this API are available for routing. Requests that reference a non-enabled model are rejected. For full request and response schemas, see Enable model for workspace and Disable model for workspace.

Restrict Model Access by Project

Once a model is enabled, workspace admins can see an Access control icon next to it. Select it to choose which projects can use the model, in one of two modes:
  • All projects (default): every project in the workspace can use the model.
  • Custom: each project gets its own on/off toggle, letting admins grant or revoke access per project.
Changes save immediately; there is no separate save action.
Access control is admin-only. Members with other roles see only the models an admin has approved within their chosen projects.

Onboarding Private Models

Onboard private models by choosing Model at the top-right of the screen. This is useful when hosting a fine-tuned model or any model deployed on a private provider such as an OpenAI-compatible endpoint, Azure AI Foundry, or AWS Bedrock.

Private Models Providers

Referencing Private Models in Code

When referencing private models through the SDKs, API, or Supported Libraries, the model is referenced by the following string: <workspacename>@<provider>/<modelname>.
Example: corp@azure/gpt-5.6-sol

Bring Your Own Key (BYOK)

To start using models, connect provider API keys via BYOK in the AI Gateway sidebar.