Skip to main content

Available models in Unity Gateway

Databricks hosts the latest partner and open-weight models. Use Unity Gateway to access these models for applications, coding agents, and AI workloads on your data.

Start querying models.

Models available through Databricks​

The following models are hosted by Databricks. Each model name links to its specifications.

Model assets in system.ai are listed globally. Seeing a model there does not mean it is available in your workspace. Availability depends on your workspace region, cross-geo settings, and model availability. Check the Unity Gateway UI for the corresponding model service and confirm that you have permission to query it.

Listing a model does not send customer data to it for inference. See Model availability by region for regional details.

For deprecated models and recommended replacements, see Deprecated and retired models.

Use a model​

Your task

How to use the models

Build an application or agent

Call a model through the Unity Gateway API. System-provided model services in system.ai let you start without creating a custom model service. See Query models.

Apply AI to your data

Use ai_query to send prompts to a supported model from SQL, notebooks, or batch pipelines. See Use ai_query.

Use a coding agent

Connect your coding agent to models through Unity Gateway. See Supported coding agents.

Your task

How to use the models

Build an application or agent

Call a model through the Unity Gateway API. System-provided model services in system.ai let you start without creating a custom model service. See Query models.

Apply AI to your data

Use ai_query to send prompts to a supported model from SQL, notebooks, or batch pipelines. See Use ai_query.

Use a coding agent

Connect your coding agent to models through Unity Gateway. See Supported coding agents.

Supported models and model identifiers vary by interface. See the linked guides for requirements and examples.

System-provided model services ​

Databricks provides ready-to-use, pay-per-token model services in system.ai, such as system.ai.claude-opus-5. New services are added as foundation models become available.

  • A system user owns them, and you cannot delete them.
  • By default, only metastore administrators can modify them. A metastore administrator can delegate management by granting the MANAGE privilege.

To restrict access, see Discover and govern access to model services.

Choose serving capacity​

Option

Recommended for

How it works

Pay-per-token

Workloads with variable or intermittent usage.

Pay for the tokens you use without provisioning dedicated serving capacity.

On-demand provisioned throughput

Workloads that need dedicated capacity with flexibility to adjust it as requirements change.

Pay for allocated capacity with no term commitment.

Reserved provisioned throughput

Business-critical applications or agents with predictable usage and sustained capacity needs.

Reserve dedicated capacity for a fixed term and pay for the full reservation, regardless of usage.

Option

Recommended for

How it works

Pay-per-token

Workloads with variable or intermittent usage.

Pay for the tokens you use without provisioning dedicated serving capacity.

On-demand provisioned throughput

Workloads that need dedicated capacity with flexibility to adjust it as requirements change.

Pay for allocated capacity with no term commitment.

Reserved provisioned throughput

Business-critical applications or agents with predictable usage and sustained capacity needs.

Reserve dedicated capacity for a fixed term and pay for the full reservation, regardless of usage.

Unity Gateway model services can route to provisioned-throughput destinations. See each option's documentation for eligible models and setup requirements.

For batch inference with ai_query, use the AI Functions guidance to select supported models and serving options. Model and regional support vary by serving option.

Connect other model providers​

You can also use Unity Gateway with your own model provider account. Connect the model provider and govern access using your organization's credentials.

See Connect an external model provider.