Databricks releases, every cloud
- Google Gemini 3 Pro Preview now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions for pay-per-token and batch inference workloads.
- OpenAI GPT-5.1 now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions for pay-per-token and batch inference workloads.
- New MCP server tab
Shows managed and external servers, and lists available servers from Databricks Marketplace.
- Function calling and structured output now available using the Alibaba Cloud Qwen3-Next Instruct model (Beta)
Supports Qwen3-Next 80B A3B Instruct model.
- Supervisor Agent now supports Unity Catalog functions and external MCP servers
Delegates tasks to Unity Catalog functions and external MCP servers.
- Gemini 2.5 Pro and Flash models now supported on pay-per-token endpoints
Supports pay-per-token endpoints and Function calling.
- Run the change, do not just read it. Hands-on labs in your own Databricks workspace, graded when you submit. Browse labs
- Improved autoscaling behavior for Mosaic AI Model Serving
Ignores brief traffic surges, scaling on sustained load increases. Reduces unnecessary concurrency scaling.
- SQL MCP server now available (Beta)
Executes SQL queries against Unity Catalog tables using SQL warehouses. Available at https://<workspace-hostname>/api/2.0/mcp/sql.
- Mosaic AI Model Serving now supports OpenAI GPT 5 models (preview)
Supports GPT-5, GPT-5 mini, and GPT-5 nano models for batch inference with AI Functions.
- Anthropic Claude Sonnet 4.5 now available as a Databricks-hosted model
Accessible via Foundation Model APIs on a pay-per-token basis.
- Prompt caching is now supported for Claude models
Caches thinking messages, images, and tool use via cache_control parameter.
- Mosaic AI Agent Framework supports automatic authentication passthrough for Lakebase resources
Requires MLFlow 3.3.2 or above.
- Token-based rate limits now available on AI Gateway
Configures rate limits on model serving endpoints, replacing legacy governance settings, and requires configuration updates.
- OpenAI GPT OSS models now support function calling
Supports function and tool calling on Databricks-hosted models.
- OpenAI GPT OSS models now supported on provisioned throughput
Supports GPT OSS 120B and 20B models on provisioned throughput.
- Schedule jobs for Serverless GPU compute workloads
Creates and schedules Serverless GPU jobs via notebook scheduling.
- Use OpenAI GPT OSS models for batch inference and structured outputs
Optimized for AI Functions with batch inference and structured outputs. Supports ai_query().
- OpenAI GPT OSS models now available on Mosaic AI Model Serving
Supports GPT OSS 120B and 20B as hosted foundation models, accessible via pay-per-token APIs.
- Get this table by email. One Monday mail covering the week, filtered to the clouds and products you run. Weekly digest
Headlines, dates and product areas are Databricks' own, and every item links to the note it came from. The one-line summaries are ours.