Your Data Lakehouse Magazine and Community
News, newsletters, and a thriving community for everyone building on the data lakehouse ecosystem. Stay informed with weekly roundups, deep dives into Apache Iceberg, and expert guides from practitioners.
- 500+
- Articles & tutorials
- 300+
- Glossary entries
- 5
- Free lakehouse books
- Weekly
- Newsletter & roundups

Find your way around
Four ways into the Hub — reference, reading, watching, and meeting people.
Knowledge Base
A searchable glossary of lakehouse, Iceberg, and catalog terminology.
BrowseThe Blog
Tutorials, architecture deep dives, and ecosystem news, published weekly.
BrowseVideo Explainers
Short, focused walkthroughs of the concepts that are hard to read about.
BrowseEvents Calendar
Meetups, webinars, and Lakehouse Linkups happening across the community.
BrowseLatest Posts

Data Quality Tooling Compared: Great Expectations, Soda, dbt Tests, and Anomaly Detection
A comparison of Great Expectations, Soda, dbt tests, and anomaly detection, and a layered design that uses each where it fits.
Read more
The Data Team of the Agentic Era: Generalists Owning End-to-End Workflows
The case for generalists owning end-to-end data workflows with agents, the counterargument, and how to make the transition work.
Read more
dbt on Iceberg: Incremental Models on Open Tables
How dbt incremental materializations map to Iceberg operations, and the configuration, predicates, and maintenance that keep them healthy.
Read moreMust Read Articles
Deep dives into Apache Iceberg, agentic AI, and modern data lakehouse architecture by Alex Merced.
What is Apache Iceberg? The Table Format Revolution
Learn how Apache Iceberg turned raw Parquet files in S3 into a fully ACID-compliant, time-traveling analytical database without moving your data out of object storage.
Read articleAgentic Analytics on the Apache Lakehouse
How autonomous AI agents replace manual dashboard querying by reading governed semantic layers directly on your data lakehouse.
Read articleEcosystemThe 2025 State of the Apache Iceberg Ecosystem
Survey data from data professionals on adoption rates, popular tooling, and where the ecosystem is heading through 2026.
Read articleData LakehouseWhat Are Table Formats and Why Were They Needed?
Table formats solved the ACID, schema evolution, and query performance problems that turned data lakes into unmanageable swamps.
Read articleApache Iceberg2026 Intro to Apache Iceberg
A beginner-to-intermediate introduction covering the metadata layer, catalog integrations, and why Iceberg became the dominant table format.
Read articleAgentic AIWhy AI Fails Without a Semantic Layer
The business context and governed metrics AI agents need to generate accurate, trustworthy analytical answers at scale.
Read articleWhy Build a Data Lakehouse?
Unified Access
Eliminate data silos. Query your data where it lives—in S3, ADLS, or GCS—without moving it.
Open Standards
Avoid vendor lock-in. Use open formats like Apache Iceberg and open catalogs to keep your data accessible to any engine.
High Performance
Achieve sub-second query performance on data lake scale datasets using engines like Dremio.
The Modern Data Stack is Open
The Data Lakehouse Hub is your central resource for tutorials, architectural guides, and community support. Whether you're migrating from a warehouse or building from scratch, we have the resources to help you succeed with open data standards.
Meet the author
Data Lakehouse Reading
Full-length books on Apache Iceberg, Apache Polaris, and lakehouse architecture — no paywall.

Stay Ahead of the Curve
Subscribe to our newsletter and event calendar to get the latest tutorials, webinars, and meetups delivered to your inbox.




