{"product_id":"engineering-lakehouses-with-open-table-formats-build-scalable-and-efficient-lakehouses-with-apache-iceberg-apache-hudi-and-delta-lake-paperback","title":"Engineering Lakehouses with Open Table Formats: Build scalable and efficient lakehouses with Apache Iceberg, Apache Hudi, and Delta Lake - Paperback","description":"\u003cdiv\u003e\u003cp style=\"text-align: right;\"\u003e\u003ca href=\"https:\/\/reportcopyrightinfringement.com\/\" target=\"_blank\" rel=\"nofollow\"\u003e\u003cb\u003eReport copyright infringement\u003c\/b\u003e\u003c\/a\u003e\u003c\/p\u003e\u003c\/div\u003e\u003cp\u003eby \u003cb\u003eDipankar Mazumdar\u003c\/b\u003e (Author), \u003cb\u003eVinoth Govindarajan\u003c\/b\u003e (Author)\u003c\/p\u003e\u003cp\u003e\u003c\/p\u003e\u003cp\u003e\u003cstrong\u003eJump-start your journey toward mastering open data architectural patterns by learning the fundamentals and applications of open table formats\u003c\/strong\u003e\u003c\/p\u003e\u003cp\u003e\u003cstrong\u003eKey Features: \u003c\/strong\u003e\u003c\/p\u003e\u003cp\u003e- Build lakehouses with open table formats using compute engines such as Apache Spark, Flink, Trino, and Python\u003c\/p\u003e\u003cp\u003e- Optimize lakehouses with techniques such as pruning, partitioning, compaction, indexing, and clustering\u003c\/p\u003e\u003cp\u003e- Find out how to enable seamless integration, data management, and interoperability using Apache XTable\u003c\/p\u003e\u003cp\u003e- Purchase of the print or Kindle book includes a free PDF eBook\u003c\/p\u003e\u003cp\u003e\u003cstrong\u003eBook Description: \u003c\/strong\u003e\u003c\/p\u003e\u003cp\u003eEngineering Lakehouses with Open Table Formats provides detailed insights into lakehouse concepts, and dives deep into the practical implementation of open table formats such as Apache Iceberg, Apache Hudi, and Delta Lake.\u003c\/p\u003e\u003cp\u003eYou'll explore the internals of a table format and learn in detail about the transactional capabilities of lakehouses. You'll also get hands on with each table format with exercises using popular computing engines, such as Apache Spark, Flink, Trino, and Python-based tools. The book addresses advanced topics, including performance optimization techniques and interoperability among different formats, equipping you to build production-ready lakehouses. With step-by-step explanations, you'll get to grips with the key components of lakehouse architecture and learn how to build, maintain, and optimize them.\u003c\/p\u003e\u003cp\u003eBy the end of this book, you'll be proficient in evaluating and implementing open table formats, optimizing lakehouse performance, and applying these concepts to real-world scenarios, ensuring you make informed decisions in selecting the right architecture for your organization's data needs.\u003c\/p\u003e\u003cp\u003e\u003cstrong\u003eWhat You Will Learn: \u003c\/strong\u003e\u003c\/p\u003e\u003cp\u003e- Explore lakehouse fundamentals, such as table formats, file formats, compute engines, and catalogs\u003c\/p\u003e\u003cp\u003e- Gain a complete understanding of data lifecycle management in lakehouses\u003c\/p\u003e\u003cp\u003e- Learn how to systematically evaluate and choose the right lakehouse table format\u003c\/p\u003e\u003cp\u003e- Optimize performance with sorting, clustering, and indexing techniques\u003c\/p\u003e\u003cp\u003e- Use the open table format data with ML frameworks like TensorFlow and MLflow\u003c\/p\u003e\u003cp\u003e- Interoperate across different table formats with Apache XTable and UniForm\u003c\/p\u003e\u003cp\u003e- Secure your lakehouse with access controls and ensure regulatory compliance\u003c\/p\u003e\u003cp\u003e\u003cstrong\u003eWho this book is for: \u003c\/strong\u003e\u003c\/p\u003e\u003cp\u003eThis book is for data engineers, software engineers, and data architects who want to deepen their understanding of open table formats, such as Apache Iceberg, Apache Hudi, and Delta Lake, and see how they are used to build lakehouses. It is also valuable for professionals working with traditional data warehouses, relational databases, and data lakes who wish to transition to an open data architectural pattern. Basic knowledge of databases, Python, Apache Spark, Java, and SQL is recommended for a smooth learning experience.\u003c\/p\u003e\u003cp\u003e\u003cstrong\u003eTable of Contents\u003c\/strong\u003e\u003c\/p\u003e\u003cp\u003e- Open Data Lakehouse: A New Architectural Paradigm\u003c\/p\u003e\u003cp\u003e- Transactional Capabilities of the Lakehouse\u003c\/p\u003e\u003cp\u003e- Apache Iceberg Deep Dive\u003c\/p\u003e\u003cp\u003e- Apache Hudi Deep Dive\u003c\/p\u003e\u003cp\u003e- Delta Lake Deep Dive\u003c\/p\u003e\u003cp\u003e- Catalog and Metadata Management\u003c\/p\u003e\u003cp\u003e- Interoperability in Lakehouses\u003c\/p\u003e\u003cp\u003e- Performance Optimization and Tuning in a Lakehouse\u003c\/p\u003e\u003cp\u003e- Data Governance and Security in Lakehouses\u003c\/p\u003e\u003cp\u003e- Evaluating and Selecting Open Table Formats\u003c\/p\u003e\u003cp\u003e- Real-World Applications and Learnings\u003c\/p\u003e\n            \u003cdiv\u003e\n\u003cstrong\u003eNumber of Pages:\u003c\/strong\u003e 414\u003c\/div\u003e\n            \u003cdiv\u003e\n\u003cstrong\u003eDimensions:\u003c\/strong\u003e 0.85 x 9.25 x 7.5 IN\u003c\/div\u003e\n            \u003cdiv\u003e\n\u003cstrong\u003ePublication Date:\u003c\/strong\u003e December 26, 2025\u003c\/div\u003e\n            ","brand":"Books by splitShops","offers":[{"title":"Default Title","offer_id":47510744465586,"sku":"9781836207238","price":64.78,"currency_code":"USD","in_stock":true}],"thumbnail_url":"\/\/cdn.shopify.com\/s\/files\/1\/0770\/3891\/1666\/files\/uOKJFjd99H9781836207238.webp?v=1779650641","url":"https:\/\/box.dadyminds.org\/products\/engineering-lakehouses-with-open-table-formats-build-scalable-and-efficient-lakehouses-with-apache-iceberg-apache-hudi-and-delta-lake-paperback","provider":"DADYMINDS BOX","version":"1.0","type":"link"}