7:50 · Music video
The Iceberg Open Lakehouse
From Parquet files and snapshots to catalogs, engines, governance, and AI-ready semantics.
Open videoNews, newsletters, and a thriving community for everyone building on the data lakehouse ecosystem. Stay informed with weekly roundups, deep dives into Apache Iceberg, and expert guides from practitioners.

Featured music video
A beat-synced, scene-rich guide to Apache Iceberg and the open architecture growing around it.
7:50 · Music video
From Parquet files and snapshots to catalogs, engines, governance, and AI-ready semantics.
Open videoFour ways into the Hub — reference, reading, watching, and meeting people.
A searchable glossary of lakehouse, Iceberg, and catalog terminology.
BrowseTutorials, architecture deep dives, and ecosystem news, published weekly.
BrowseShort, focused walkthroughs of the concepts that are hard to read about.
BrowseMeetups, webinars, and Lakehouse Linkups happening across the community.
Browse
A comparison of Great Expectations, Soda, dbt tests, and anomaly detection, and a layered design that uses each where it fits.
Read more
The case for generalists owning end-to-end data workflows with agents, the counterargument, and how to make the transition work.
Read more
How dbt incremental materializations map to Iceberg operations, and the configuration, predicates, and maintenance that keep them healthy.
Read moreDeep dives into Apache Iceberg, agentic AI, and modern data lakehouse architecture by Alex Merced.
Learn how Apache Iceberg turned raw Parquet files in S3 into a fully ACID-compliant, time-traveling analytical database without moving your data out of object storage.
Read articleHow autonomous AI agents replace manual dashboard querying by reading governed semantic layers directly on your data lakehouse.
Read articleEcosystemSurvey data from data professionals on adoption rates, popular tooling, and where the ecosystem is heading through 2026.
Read articleData LakehouseTable formats solved the ACID, schema evolution, and query performance problems that turned data lakes into unmanageable swamps.
Read articleApache IcebergA beginner-to-intermediate introduction covering the metadata layer, catalog integrations, and why Iceberg became the dominant table format.
Read articleAgentic AIThe business context and governed metrics AI agents need to generate accurate, trustworthy analytical answers at scale.
Read articleEliminate data silos. Query your data where it lives—in S3, ADLS, or GCS—without moving it.
Avoid vendor lock-in. Use open formats like Apache Iceberg and open catalogs to keep your data accessible to any engine.
Achieve sub-second query performance on data lake scale datasets using engines like Dremio.
The Data Lakehouse Hub is your central resource for tutorials, architectural guides, and community support. Whether you're migrating from a warehouse or building from scratch, we have the resources to help you succeed with open data standards.
Meet the author
Full-length books on Apache Iceberg, Apache Polaris, and lakehouse architecture — no paywall.

Subscribe to our newsletter and event calendar to get the latest tutorials, webinars, and meetups delivered to your inbox.