Databricks releases, every cloud
- OpenAI GPT-5.4 mini and GPT-5.4 nano now available as Databricks-hosted models
Access via Foundation Model APIs pay-per-token, requires compliance with OpenAI's Acceptable Use Policy.
- Empty vector search endpoints no longer incur charges
Charges stop 24 hours after last index is deleted.
AWS, read it on their docsAzure, read it on their docsGCP, read it on their docsSAP(not on this cloud) - Google Gemini 3.1 Flash Lite now available as a Databricks-hosted model
Access via Foundation Model APIs or AI Functions, supporting pay-per-token and batch inference workloads.
AWS, read it on their docsAzure, read it on their docsGCP, read it on their docsSAP(not on this cloud) - OpenAI GPT-5.4 now available as a Databricks-hosted model
Accessible via Foundation Model APIs on a pay-per-token basis, requires compliance with OpenAI's Acceptable Use Policy.
- Managed OAuth flows for external MCP servers
Eliminates need to register own OAuth app, supports Glean, GitHub, Google Drive, and SharePoint APIs.
AWS, read it on their docsAzure, read it on their docsGCP, read it on their docsSAP(not on this cloud) - OpenAI GPT-5.3 Codex now available as a Databricks-hosted model
Available via Foundation Model APIs pay-per-token, requiring compliance with OpenAI's Acceptable Use Policy.
- Run the change, do not just read it. Hands-on labs in your own Databricks workspace, graded when you submit. Browse labs
- Change to use_case field in VPC endpoint API responses Coming soon
Changes from WORKSPACE_ACCESS to GENERAL_ACCESS, see API docs for details.
- Google Gemini 3.1 Pro Preview now available as a Databricks-hosted model
Access via Foundation Model APIs, AI Functions for batch inference.
- Anthropic Claude Sonnet 4.6 now available as a Databricks-hosted model
Accessible via Foundation Model APIs on a pay-per-token basis.
- Qwen3-Embedding-0.6B now available in Public Preview as a Databricks-hosted model
Accessible via Foundation Model APIs on a pay-per-token basis.
- Anthropic Claude Opus 4.6 now available as a Databricks-hosted model
Accessible via Foundation Model APIs, supports pay-per-token and query reasoning and vision models.
- OpenAI GPT-5.2 Codex now available as a Databricks-hosted model
Excels at code generation and refactoring tasks. Accessible via Foundation Model APIs pay-per-token.
- OpenAI GPT-5.2 Codex now available as a Databricks-hosted model
Excels at code generation and debugging tasks, accessed via Foundation Model APIs pay-per-token.
- Managed MCP servers is now Public Preview
Allows AI agents to connect to Databricks resources and external APIs.
- OpenAI GPT-5.1 Codex Max and Codex Mini now available as Databricks-hosted models
Access via Foundation Model APIs pay-per-token, subject to OpenAI's Acceptable Use Policy.
- OpenAI GPT-5.1 Codex Max and Codex Mini now available as Databricks-hosted models
Supports code generation and refactoring via Foundation Model APIs. Requires pay-per-token access.
- Anthropic Claude Haiku 4.5 now available as a Databricks-hosted model
Accessible via Foundation Model APIs on a pay-per-token basis.
- Gemini 3 Flash now available as a Databricks-hosted model
Offers speed and scale for video analysis and data extraction. Requires no setup.
- Get this table by email. One Monday mail covering the week, filtered to the clouds and products you run. Weekly digest
- OpenAI GPT-5.2 now available as a Databricks-hosted model
Accessible via Foundation Model APIs on a pay-per-token basis.
- Anthropic Claude Opus 4.5 now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions, supports pay-per-token and batch inference workloads.
- Google Gemini 3 Pro Preview now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions, supporting pay-per-token and batch inference workloads.
- Google Gemini 3 Pro Preview now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions for pay-per-token and batch inference workloads.
- OpenAI GPT-5.1 now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions for pay-per-token and batch inference workloads.
- New MCP server tab
Shows managed and external servers, and lists available servers from Databricks Marketplace.
- Feedback model deprecated for AI agents
Replaced by log_feedback API in MLflow 3.
- Request logs and assessment logs tables deprecated
Replaced by MLflow 3 for real-time tracing, stops populating November 15, 2025.
- OpenAI GPT 5 models are now generally available in Mosaic AI Model Serving
Includes GPT-5 mini and nano models.
- Gemini 2.5 Pro and Flash models now supported on pay-per-token endpoints
Supports pay-per-token endpoints and Function calling.
- Improved autoscaling behavior for Mosaic AI Model Serving
Ignores brief traffic surges, scaling on sustained load increases. Reduces unnecessary concurrency scaling.
- SQL MCP server now available (Beta)
Executes SQL queries against Unity Catalog tables using SQL warehouses. Available at https://<workspace-hostname>/api/2.0/mcp/sql.
- Mosaic AI Model Serving now supports OpenAI GPT 5 models (preview)
Supports GPT-5, GPT-5 mini, and GPT-5 nano models for batch inference with AI Functions.
- Anthropic Claude Sonnet 4.5 now available as a Databricks-hosted model
Accessible via Foundation Model APIs on a pay-per-token basis.
- Prompt caching is now supported for Claude models
Caches thinking messages, images, and tool use via cache_control parameter.
- AI agents: Authorize on-behalf-of-user Public Preview
Lets agents act as the Databricks user who runs the query for added security.
- External MCP servers are in Beta
Enables agent access to external tools, requires connecting to Model Context Protocol servers.
- Token-based rate limits now available on AI Gateway
Configures rate limits on model serving endpoints, replacing legacy governance settings, and requires configuration updates.
- OpenAI GPT OSS models now support function calling
Supports function and tool calling on Databricks-hosted models.
- OpenAI GPT OSS models now supported on provisioned throughput
Supports GPT OSS 120B and 20B models on provisioned throughput.
- Use OpenAI GPT OSS models for batch inference and structured outputs
Optimized for AI Functions with batch inference and structured outputs. Supports ai_query().
- OpenAI GPT OSS models now available on Mosaic AI Model Serving
Supports GPT OSS 120B and 20B as hosted foundation models, accessible via pay-per-token APIs.
Headlines, dates and product areas are Databricks' own, and every item links to the note it came from. The one-line summaries are ours.