# Advanced Spark (Day 2 Lab)

> In this lab, Zach shows the students how to use Glue Job Runner and Iceberg to optimize the data processing. He goes over setting up the job, running Python functions, and using UDFs. He also demonstrates how to monitor the job and view the output table. Plus, he explains the benefits of using Iceb…

- Web page: https://www.dataexpert.io/lesson/spark-batch-day-2-lab-v4
- Program: [Data Engineering Mastery Course](https://www.dataexpert.io/program/data-engineering-mastery-course)
- Module: Week 4: Batch Pipelines with Apache Spark
- Access: Requires enrollment in Data Engineering Mastery Course
- Length: 37 min video
- Skills: Python, Apache Spark
- Academy: DataExpert.io Academy

## About this lesson

In this lab, Zach shows the students how to use Glue Job Runner and Iceberg to optimize the data processing. He goes over setting up the job, running Python functions, and using UDFs. He also demonstrates how to monitor the job and view the output table. Plus, he explains the benefits of using Iceberg for data compression and partitioning. [Recorded on May30th, 2024]

## Navigation

- Previous lesson: [Spark Data Quality (Day 3 Lab)](https://www.dataexpert.io/lesson/spark-batch-day-3-lab-v4.md)
- Next lesson: [Spark Data Quality (Day 3 Lecture)](https://www.dataexpert.io/lesson/spark-batch-day-3-lecture-v4.md)
