Databricks releases, every cloud
- Protobuf tensor input for custom model serving endpoints (Public Preview)
Replaces JSON with serialized KServe v2 ModelInferRequest. Requires endpoints deployed after July 9, 2026.
AWS, read it on their docsAzure, read it on their docsGCP, read it on their docsSAP(not on this cloud) - Custom Docker images for AI Runtime CLI workloads (Beta)
Allows custom system library versions and complex dependencies, requires setup via Use custom Docker images.
- AI Runtime CLI (Beta)
Submits and manages distributed training workloads on serverless GPU compute from a local machine.
- Serve custom LLMs with Custom Model Serving (Beta)
Supports multimodal models and PEFT recipes, unlike Foundation Model APIs.
- AI Runtime 1xH100 accelerator (Beta)
Supports 1xH100 accelerator, see Hardware options for details.
- AI Runtime is now in Public Preview
Adds GPU support to serverless compute for deep learning workloads.
- Run the change, do not just read it. Hands-on labs in your own Databricks workspace, graded when you submit. Browse labs
- Endpoint telemetry for custom model serving endpoints (Beta)
Stores logs, traces, and metrics in Unity Catalog Delta tables using OpenTelemetry.
Headlines, dates and product areas are Databricks' own, from the release notes published for each cloud, and every item links to the note it came from. The one-line summaries are ours. Not a Databricks product and not affiliated with Databricks, Inc.