Databricks releases, every cloud
- Anthropic Claude Opus 4.5 now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions, supports pay-per-token and batch inference workloads.
- Google Gemini 3 Pro Preview now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions, supporting pay-per-token and batch inference workloads.
- Google Gemini 3 Pro Preview now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions for pay-per-token and batch inference workloads.
- OpenAI GPT-5.1 now available as a Databricks-hosted model
Accessible via Foundation Model APIs and AI Functions for pay-per-token and batch inference workloads.
- New MCP server tab
Shows managed and external servers, and lists available servers from Databricks Marketplace.
- Function calling and structured output now available using the Alibaba Cloud Qwen3-Next Instruct model (Beta)
Supports Qwen3-Next 80B A3B Instruct model.
- Run the change, do not just read it. Hands-on labs in your own Databricks workspace, graded when you submit. Browse labs
- Supervisor Agent now supports Unity Catalog functions and external MCP servers
Delegates tasks to Unity Catalog functions and external MCP servers.
- Gemini 2.5 Pro and Flash models now supported on pay-per-token endpoints
Supports pay-per-token endpoints and Function calling.
- Improved autoscaling behavior for Mosaic AI Model Serving
Ignores brief traffic surges, scaling on sustained load increases. Reduces unnecessary concurrency scaling.
- SQL MCP server now available (Beta)
Executes SQL queries against Unity Catalog tables using SQL warehouses. Available at https://<workspace-hostname>/api/2.0/mcp/sql.
- Mosaic AI Model Serving now supports OpenAI GPT 5 models (preview)
Supports GPT-5, GPT-5 mini, and GPT-5 nano models for batch inference with AI Functions.
- Anthropic Claude Sonnet 4.5 now available as a Databricks-hosted model
Accessible via Foundation Model APIs on a pay-per-token basis.
- Prompt caching is now supported for Claude models
Caches thinking messages, images, and tool use via cache_control parameter.
- Mosaic AI Agent Framework supports automatic authentication passthrough for Lakebase resources
Requires MLFlow 3.3.2 or above.
- Token-based rate limits now available on AI Gateway
Configures rate limits on model serving endpoints, replacing legacy governance settings, and requires configuration updates.
- OpenAI GPT OSS models now support function calling
Supports function and tool calling on Databricks-hosted models.
- OpenAI GPT OSS models now supported on provisioned throughput
Supports GPT OSS 120B and 20B models on provisioned throughput.
- Schedule jobs for Serverless GPU compute workloads
Creates and schedules Serverless GPU jobs via notebook scheduling.
- Get this table by email. One Monday mail covering the week, filtered to the clouds and products you run. Weekly digest
- Use OpenAI GPT OSS models for batch inference and structured outputs
Optimized for AI Functions with batch inference and structured outputs. Supports ai_query().
- OpenAI GPT OSS models now available on Mosaic AI Model Serving
Supports GPT OSS 120B and 20B as hosted foundation models, accessible via pay-per-token APIs.
Headlines, dates and product areas are Databricks' own, and every item links to the note it came from. The one-line summaries are ours.