Курс от CourseraLearn how to execute, monitor, and validate data pipelines through this practical course in the Data Engineering Skill Path. You will develop essential competencies, including running pre-defined pipelines, monitoring execution logs for successful completion, identifying and investigating failures, executing routine data loading jobs, and applying predefined data quality checks to ensure accurate and reliable warehouse data. Through hands-on work with Apache Airflow, Bash-based ETL concepts, and warehouse validation techniques, you will gain confidence operating real-world data workflows. This course integrates expertise from multiple IBM instructional teams, offering a multi-perspective view of pipeline orchestration, ETL operations, and warehouse validation. You will progress from visualizing and debugging DAG-based workflows in Airflow, to understanding pipeline behavior in batch and streaming contexts, and finally to applying quality checks and verification methods within a modern data warehouse environment. The curriculum combines conceptual understanding with practical exercises, preparing you to maintain, troubleshoot, and validate pipelines in production-like settings. Ideal for aspiring data engineers and learners seeking strong foundational skills in pipeline execution, monitoring, and operational data quality assurance.
5 модулей · 46 учебных материалов

Преподаватель курса