This comprehensive 3-day training focuses on Apache Iceberg, a high-performance table format for huge analytic datasets, designed to overcome the limitations of traditional data lakes. Participants will start by understanding the need for modern table formats and exploring Iceberg’s architecture, core features, and capabilities through lectures and hands-on labs.
The training covers schema and partition evolution, snapshot-based isolation, ACID transactions, catalog management, and advanced query optimization. Attendees will gain practical skills in creating, querying, and maintaining Iceberg tables using engines like Spark, Flink, and Trino, as well as integrating Iceberg into Lakehouse architectures.
By the end of the training, participants will be equipped to design and manage Iceberg-based data lakes, implement incremental data processing, and apply governance and security best practices for production workloads.
Duration: 3 Days
Course Code: BDT 513
Learning Objectives:
After this training, participants will be able to:
This course is ideal for:
Training material provided: Yes (Digital format)