About
the event
The Data + AI Summit is the world's largest data, analytics, and AI conference - held annually at the Moscone Center in San Francisco. In 2026, more than 30,000 data and AI professionals are expected to gather in person, with tens of thousands more joining virtually, together representing over 160 countries.
The summit brings together Databricks' founders, leading AI builders, open-source contributors, and global business leaders for a multi-day exploration of what is next in data and AI. Attendees get a first look at Databricks' newest product innovations and can deep-dive into the latest breakthroughs across data engineering, governance, analytics, and agentic AI systems.
Event highlights and what to expect
- 800+ sessions spanning data engineering, analytics, agents, and AI
- Keynotes from Databricks founders and leading AI builders
- First looks at brand-new product innovations
- Hands-on training, instructor-led workshops, and brand-new certification courses
- Deep-dives into Delta Lake, Mosaic AI, Apache Iceberg, Spark, MLflow, Unity Catalog, Lakeflow
- Discussions on open-source frameworks like LangChain, PyTorch, and dbt
- 80+ special interest and networking events
- Virtual livestream access for global attendees
- 800+ sessions spanning data engineering, analytics, agents, and AI
- Keynotes from Databricks founders and leading AI builders
- First looks at brand-new product innovations
- Hands-on training, instructor-led workshops, and brand-new certification courses
- Deep-dives into Delta Lake, Mosaic AI, Apache Iceberg, Spark, MLflow, Unity Catalog, Lakeflow
- Discussions on open-source frameworks like LangChain, PyTorch, and dbt
- 80+ special interest and networking events
- Virtual livestream access for global attendees
Our Services
Delta Lake Architecture
Open lakehouse foundation with ACID transactions, scalable metadata handling, and time-travel capabilities for analytics and AI.
Unity Catalog Data Models
Reusable, industry-specific data models for fintech, healthtech, adtech, supply chain, and HVACR - with full governance and lineage.
Legacy ETL Migration
Migrate from brittle batch ETL and legacy warehouses to governed lakehouse architecture ready for advanced analytics and AI.
Medallion Architecture
Bronze, Silver, Gold structured data layers for systematic cleansing, validation, and business-rule enforcement at scale.
Databricks Mosaic AI
Build and deploy production AI agents on Agent Bricks. Fine-tune foundation models and orchestrate multi-step intelligent workflows.
Agent Bricks
Custom AI agents for fintech, healthtech, supply chain, and HVACR - built natively on Databricks infrastructure and ready to scale.
MLflow & Model Registry
End-to-end ML lifecycle management - experiment tracking, model versioning, and deployment pipelines built on MLflow.
Predictive Models
Industry-specific forecasting, anomaly detection, and intelligent automation models trained and served on the Databricks platform.
Databricks SQL
High-performance SQL analytics on Delta Lake - from ad-hoc exploration to production BI dashboards at enterprise scale.
BI & Reporting Layer
Connect Tableau, Power BI, and Looker to a governed Databricks lakehouse, with semantic layers that business users can trust.
Time-to-Insight Acceleration
Optimized query performance, auto-scaling compute, and pre-built domain metrics to dramatically reduce time-to-insight.
Unity Catalog Governance
Unified data governance - row-level security, column masking, and attribute-based access control across all data assets.
Data Lineage & Auditing
Full end-to-end lineage from source to model. Know exactly where your data comes from and where it flows.
Compliance Frameworks
Built-in controls for GDPR, HIPAA, SOC 2, and PCI DSS - designed for regulated industries like fintech and healthtech.
Lakewatch
Continuous data quality monitoring and anomaly detection across your lakehouse - proactively alerting on schema drift, pipeline failures, and data freshness SLA breaches.
Real-Time Streaming Pipelines
Low-latency ingestion using Databricks Structured Streaming - replacing batch ETL with live data flows into Delta Lake.
Kafka + Redis Integration
Streaming ingestion and near real-time processing via Kafka and Redis - for fraud detection, inventory, and personalization.
Databricks Workflows
Automated ETL job scheduling, orchestration, and monitoring - with GitHub integration for version-controlled deployments.
Meet our
Experts
Skip the small talk. Connect directly with the minds behind our most impactful solutions and get tailored insights for your next major project.


