Refresh Your Home For Fall
To see product details, add this item to your cart. You can always remove it later.
Shipper / Seller
Amazon.com
Amazon.com
Shipper / Seller
Amazon.com
Returns
30-day refund / replacement
30-day refund / replacement
This item can be returned in its original condition for a full refund or replacement within 30 days of receipt.
Read full return policy
Payment
Secure transaction
Your transaction is secure
We work hard to protect your security and privacy. Our payment security system encrypts your information during transmission. We don’t share your credit card details with third-party sellers, and we don’t sell your information to others. Learn more
Gift options
Available at checkout
Available at checkout This item is a gift. Change
At checkout, you can add a custom message, a gift receipt for easy returns and have the item gift-wrapped
Added to

Sorry, there was a problem.

There was an error retrieving your Wish Lists. Please try again.

Sorry, there was a problem.

List unavailable.
Kindle app logo image

Download the free Kindle app and start reading Kindle books instantly on your smartphone, tablet, or computer - no Kindle device required.

Read instantly on your browser with Kindle for Web.

Using your mobile phone camera - scan the code below and download the Kindle app.

QR code to download the Kindle App

  • Architecting an Apache Iceberg Lakehouse: A scalable, open-source data platform
  • Manning Introduces: Architecting an Apache Iceberg Lakehouse
  • VIDEO

Architecting an Apache Iceberg Lakehouse: A scalable, open-source data platform

5.0 out of 5 stars (3)

Purchase options and add-ons

Get the eBook free when you register your print book at Manning.

Design an Apache Iceberg lakehouse from scratch!

The “lakehouse” data architecture is a powerful way to combine the flexibility of data lakes with the management features of data warehouses. The open source Apache Iceberg framework delivers the scalability, reliability, and performance you want from a lakehouse without the expense and vendor lock-in of platforms like Snowflake, BigQuery, and Redshift.

In
Architecting an Apache Iceberg Data Lakehouse, data guru Alex Merced shows you:

• How to create a modular, scalable Iceberg lakehouse architecture
• Where Spark, Flink, Dremio, Polaris fit into your design
• Reliable batch and streaming ingestion pipelines
• Strategies for governance, security, and performance at scale

Apache Iceberg is an open source table format perfect for massive analytic datasets. Iceberg enables ACID transactions, schema evolution, and high-performance queries on data lakes using multiple compute engines like Spark, Trino, Flink, Presto, and Hive. An Iceberg data lakehouse enables fast, reliable analytics at scale while retaining the observability you need for compliance audits, governance, and provable data security.

Foreword by Tim Berglund. Afterword by Adi Polak.

About the technology

Apache Iceberg is an open data format that lets data lake files work like database tables. It helps turn a data lake into a more reliable and capable lakehouse.

About the book

Architecting an Apache Iceberg Lakehouse shows you how to design an open, scalable, and cost-effective lakehouse platform with Apache Iceberg. More than a set of blueprints, the book explains the reasoning behind the architecture. You’ll build a mini lakehouse by ingesting sales and marketing data from PostgreSQL into Iceberg tables with Apache Spark and then create interactive dashboards in Apache Superset. You’ll appreciate expert Alex Merced’s real-world insights about operating an Iceberg lakehouse.

What's inside

• Create a modular, scalable Iceberg lakehouse architecture
• Fit Spark, Flink, Dremio, Polaris and more into your design
• Batch and streaming ingestion pipelines
• Governance, security, and performance at scale

About the reader

For data architects familiar with the basics of a data lakehouse.

About the author

Alex Merced is Head of Developer Relations at Dremio. He shares his expertise through videos, podcasts, and articles, and leads the DataLakehouseHub.com community.

Table of Contents

Part 1
1 The world of the data lakehouse
2 Apache Iceberg and the lakehouse
3 Hands-on with Apache Iceberg
Part 2
4 Preparing for your move to Apache Iceberg
5 Selecting the storage layer
6 Architecting the ingestion layer
7 Implementing the catalog layer
8 Designing the federation layer
9 Understanding the consumption layer
Part 3 Operating your Apache Iceberg lakehouse
10 Maintaining an Iceberg lakehouse
11 Operationalizing Apache Iceberg
A The metadata tables
B Python for Apache Iceberg
C The Apache Iceberg specification

Frequently bought together

This item: Architecting an Apache Iceberg Lakehouse: A scalable, open-source data platform
$53.62
Get it as soon as Monday, Sep 21
In Stock
Ships from and sold by Amazon.com.
Total price: $00
To see our price, add these items to your cart.
Details
Added to Cart
Choose items to buy together.

Customers also bought or read

Loading...

From the Publisher

Architecting an Apache Iceberg Lakehouse header

Architecting an Apache Iceberg Lakehouse quote 1

“Gives you the practical grounding to build with confidence, and maybe even enjoy the process.”

Matt Topol Apache Iceberg PMC Member

Architecting an Apache Iceberg Lakehouse quote 2

“Building a lakehouse without this book is like building a house without a foundation.”

Roy Hasson, Microsoft

Architecting an Apache Iceberg Lakehouse quote 3

“Th e author’s passion and competence shine through in every chapter of this book.”

Joe Reis, co-author of Fundamentals of Data Engineerin

Architecting an Apache Iceberg Lakehouse cover

about the book

Architecting an Apache Iceberg Lakehouse shows you how to design a complete lakehouse architecture around Apache Iceberg, so you can make solid decisions about storage, ingestion, querying, and governance instead of just stitching tools together.

You’ll get hands-on experience building and running an Iceberg-based platform, which helps you move from theory to a working system you can adapt for production.

By the end, you’ll understand the trade-offs behind real-world data architectures and how to build a scalable, maintainable platform that supports reliable analytics at large scale.

about manning

about the authors

Manning helps developers and tech professionals stay ahead in a fast-moving industry with expert-led books, videos, and projects. Learning never stops, but it’s hard to keep up, so we focus on content that’s practical, clear, and trusted. As an independent publisher, we adapt quickly, from pioneering early-access books to offering DRM-free eBooks. Our series, like "In Action" and "In a Month of Lunches", reflect a commitment to making complex topics accessible.

Apache Kafka in Action: From basics to production
Data Pipelines with Apache Airflow, Second Edition: Orchestration for data an...
Data Pipelines with Apache Airflow
Kafka Streams in Action, Second Edition: Event-driven applications and micros...
Grokking Data Structures
Grokking Algorithms, Second Edition
Customer Reviews
4.7 out of 5 stars 5
4.6 out of 5 stars 6
4.4 out of 5 stars 77
3.8 out of 5 stars 7
4.5 out of 5 stars 18
4.7 out of 5 stars 282
User experience level Intermediate Intermediate Intermediate Intermediate Beginner Beginner
About the reader For IT operators, software architects and developers. For data engineers, machine learning engineers, DevOps, and sysadmins with intermediate Python skills. For DevOps, data engineers, machine learning engineers, and sysadmins For Java developers. For readers who know the basics of Python. No advanced math or programming skills required.
Special features Includes liveBook with out built-in AI assistant. Includes liveBook with out built-in AI assistant. Includes liveBook with out built-in AI assistant. Includes liveBook with out built-in AI assistant. Includes liveBook with out built-in AI assistant. Includes liveBook with out built-in AI assistant.
Page count 368 512 480 504 280 320

Editorial Reviews

About the Author

Alex Merced is Head of Developer Relations at Dremio, where he helps developers navigate modern data architectures. He shares his expertise through videos, podcasts, and articles, and leads the DataLakehouseHub.com community. He is the co-author of Apache Iceberg: The Definitive Guide.

Product details

  • Publisher ‏ : ‎ Manning Publications
  • Publication date ‏ : ‎ May 19, 2026
  • Language ‏ : ‎ English
  • Print length ‏ : ‎ 408 pages
  • ISBN-10 ‏ : ‎ 1633435105
  • ISBN-13 ‏ : ‎ 978-1633435100
  • Item Weight ‏ : ‎ 1.62 pounds
  • Dimensions ‏ : ‎ 7.38 x 1.02 x 9.25 inches
  • Best Sellers Rank: #1,403,146 in Books (See Top 100 in Books)
  • Customer Reviews:
    5.0 out of 5 stars (3)