Available models in Unity Gateway
Databricks hosts the latest partner and open-weight models. Use Unity Gateway to access these models for applications, coding agents, and AI workloads on your data.
Models available through Databricks
The following models are hosted by Databricks. Each model name links to its specifications.
Model assets in system.ai are listed globally. Seeing a model there does not mean it is available in your workspace. Availability depends on your workspace region, cross-geo settings, and model availability. Check the Unity Gateway UI for the corresponding model service and confirm that you have permission to query it.
Listing a model does not send customer data to it for inference. See Model availability by region for regional details.
Provider | Model family | Models |
|---|---|---|
Anthropic | Claude Haiku | |
Anthropic | Claude Sonnet | Claude Sonnet 5.5, Claude Sonnet 5, Claude Sonnet 4.6, Claude Sonnet 4.5 |
Anthropic | Claude Fable | |
Anthropic | Claude Opus | Claude Opus 5, Claude Opus 4.8, Claude Opus 4.7, Claude Opus 4.6, Claude Opus 4.5, Claude Opus 4.1 |
OpenAI | GPT | GPT-6 Astra, GPT-5.6 Sol, GPT-5.6 Terra, GPT-5.6 Luna, GPT-5.5 Pro, GPT-5.5, GPT-5.4, GPT-5.4 mini, GPT-5.4 nano, GPT-5.2, GPT-5.1, GPT-5, GPT-5 mini, GPT-5 nano |
OpenAI | GPT Codex | |
OpenAI | GPT OSS | |
Gemini Flash | Gemini 3.8 Flash, Gemini 3.7 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash, Gemini 3.5 Flash Lite, Gemini 3.1 Flash Lite, Gemini 3 Flash | |
Gemini Pro | ||
Gemini image generation | ||
Gemma | ||
Meta | Llama | Llama 4 Maverick, Llama 3.3 70B Instruct, Llama 3.1 8B Instruct |
Alibaba Cloud | Qwen | OpenJev (Qwen3.5 4B), Qwen3.5 122B A10B, Qwen3-Embedding-0.6B, Qwen3-Next 80B A3B Instruct |
xAI | Grok | |
Moonshot AI | Kimi | |
Zhipu AI | GLM | |
DeepSeek | DeepSeek | |
Thinking Machine Labs | Inkling | |
Alibaba | GTE |
For deprecated models and recommended replacements, see Deprecated and retired models.
Use a model
Your task | How to use the models |
|---|---|
Build an application or agent | Call a model through the Unity Gateway API. System-provided model services in |
Apply AI to your data | Use |
Use a coding agent | Connect your coding agent to models through Unity Gateway. See Supported coding agents. |
Supported models and model identifiers vary by interface. See the linked guides for requirements and examples.
System-provided model services
Databricks provides ready-to-use, pay-per-token model services in system.ai, such as system.ai.claude-opus-5. New services are added as foundation models become available.
- A system user owns them, and you cannot delete them.
- By default, only metastore administrators can modify them. A metastore administrator can delegate management by granting the
MANAGEprivilege.
To restrict access, see Discover and govern access to model services.
Choose serving capacity
Option | Recommended for | How it works |
|---|---|---|
Pay-per-token | Workloads with variable or intermittent usage. | Pay for the tokens you use without provisioning dedicated serving capacity. |
Workloads that need dedicated capacity with flexibility to adjust it as requirements change. | Pay for allocated capacity with no term commitment. | |
Business-critical applications or agents with predictable usage and sustained capacity needs. | Reserve dedicated capacity for a fixed term and pay for the full reservation, regardless of usage. |
Unity Gateway model services can route to provisioned-throughput destinations. See each option's documentation for eligible models and setup requirements.
For batch inference with ai_query, use the AI Functions guidance to select supported models and serving options. Model and regional support vary by serving option.
Connect other model providers
You can also use Unity Gateway with your own model provider account. Connect the model provider and govern access using your organization's credentials.
See Connect an external model provider.