К содержимому
learnspaceYOUR NEXT CHAPTER
ПРОСТРАНСТВО ОБУЧЕНИЯ
ГлавнаяКаталог курсовМоё обучениеCoursera

Знания без границ

Учитесь у лучших университетов и компаний мира.

Открыть Coursera
Интеграция
Пространство университета
Моё пространствоСтраница курса
↵
ЯЛичный кабинетСтудент
© 2026 LearnSpaceКаждый день — возможность узнать больше.Помощь
Databricks Data Engineer Associate: Practical Guide · LearnSpace
Назад в каталог
courseraПрограммирование

Databricks Data Engineer Associate: Practical Guide

Курс от Packt
Средний≈ 10.8 чАнглийский
О курсеНавыкиПрограммаПреподаватели

О курсе

This course covers all essential Databricks concepts for aspiring data engineers, including Apache Spark fundamentals, Delta Lake, performance optimization, and data pipeline creation. With real-world demonstrations, you will gain practical experience in deploying and orchestrating data engineering workflows. This course is designed to equip you with the skills needed to become a certified Databricks Data Engineer Associate. Starting with an introduction to Databricks, you will explore its role in modern data engineering and its integration with Apache Spark. From understanding basic data processing operations to building powerful data pipelines, you’ll gain hands-on experience in Spark architecture and execution, Delta Lake for data management, and the use of Databricks’ advanced features like Unity Catalog for governance and Delta Live Tables for orchestration. You’ll dive deep into Spark fundamentals, learning about data transformations, actions, and lazy evaluation, followed by the performance optimization techniques essential for data-heavy applications. The course also covers crucial data warehousing concepts like OLAP and OLTP, along with Delta Lake’s ACID transactions and time travel for data management. Through practical demonstrations, you will learn how to sign up for Databricks, create and manage notebooks, ingest and transform data, and optimize performance using partitions and parallelism. The course culminates in a comprehensive capstone project that ties everything together, giving you the experience and knowledge to pass the Databricks Certified Data Engineer Associate exam and excel in data engineering tasks. This course is designed for aspiring data engineers, data scientists, and IT professionals who wish to build a strong foundation in Databricks and Apache Spark. It's ideal for individuals looking to gain hands-on experience in data engineering workflows, Spark execution, performance optimization, and Delta Lake for managing large-scale data. Familiarity with data engineering concepts or programming basics will be helpful, but the course is suitable for both beginners and those seeking certification. The course follows a project-based approach, where you work through practical demonstrations and real-world scenarios. Each section introduces essential concepts followed by hands-on tutorials to reinforce learning. By the end, you will have mastered the skills necessary to build and optimize data pipelines and perform key tasks required for the Databricks Certified Data Engineer Associate exam. This course is based on Databricks Certified Data Engineer Associate - Practical Guide, by Yogesh Raheja, Thinknyx Technologies. This course is licensed and distributed by Packt. All rights reserved. Packt is one of the world's most prolific publishers of cutting-edge technical content. For over two decades we've made it our mission to curate and publish the knowledge of only the very best technical experts. We focus on real-world courses that help our customers get the job done, with coverage that extends across a wide range of established and cutting-edge technical topics. If you're an individual or an organisation that embraces learning by doing, Packt is the perfect fit for you.

Навыки, которые вы освоите

Apache SparkData LakesData PipelinesData WarehousingDatabricksPerformance TuningData GovernanceGitHubBusiness IntelligenceData IntegrationAnalyticsData TransformationScalabilityPySparkDashboardGit (Version Control System)Version ControlTransaction ProcessingData CleansingData Architecture

Программа курса

11 модулей · 84 учебных материалов

01Course Introduction2 материалов

Unlocking the Course Journey: Goals, Structure, and Expectations

Course IntroductionВидеоNavigating the Course IntroductionDIALOGUE
02Getting Started with Databricks8 материалов

Navigating Databricks: From Setup to Data Exploration

Section IntroductionВидео

Учитесь у экспертов

Packt - Course Instructors

Преподаватель курса

Databricks Data Engineer Associate: Practical Guide
В каталоге вашей программы

Инвестируйте в себя

Новые знания — в удобное для вас время.

Начать на Coursera

Обучение откроется на Coursera
в новой вкладке

Обучение на Coursera

≈ 10.8 ч

11 модулей

Язык: Английский

Часть программы вашего университета
What Is Databricks and Why It ExistsВидео
Demonstration: Databricks Editions and Free Signup: UI WalkthroughВидео
Demonstration: Exploring the Databricks DashboardВидео
Demonstration: Creating Your First Notebook and Ingesting DataВидео
SummaryВидео
Exploring Databricks Features and WorkflowDIALOGUE
Getting Started with DatabricksЗадание
03Understanding Spark Fundamentals11 материалов

Mastering Apache Spark: From Basics to Data Manipulation

Section IntroductionВидеоWhat Is Apache Spark & PySparkВидеоSpark Architecture (Driver, Executors, Cluster Concept)ВидеоSpark Session: The Entry Point of SparkВидеоDemonstration: Basic DataFrame OperationsВидеоDemonstration: Switching Between PySpark & SQLВидеоDemonstration: Magic Commands in Databricks & Spark SessionВидеоDemonstration: Load and Transform Data Using Apache Spark DataFramesВидеоSummaryВидеоExplaining Spark DataFrame OperationsDIALOGUEFundamentals of Apache Spark and PySparkЗадание
04Understanding Spark Execution Basics8 материалов

Decoding Spark Execution: From Plans to Performance

Section IntroductionВидеоUnderstanding Spark Plans with explain()ВидеоDemonstration: Understanding Spark Execution Plans with explain()ВидеоTransformations vs Actions & Lazy EvaluationВидеоNarrow vs Wide TransformationsВидеоSummaryВидеоAnalyzing Spark Execution PlansDIALOGUECore Concepts of Spark ExecutionЗадание
05Performance Basics7 материалов

Mastering Data Processing Efficiency in Spark

Section IntroductionВидеоPartitions and Parallelism (Conceptual)ВидеоRepartition vs CoalesceВидеоDemonstration: Repartition vs CoalesceВидеоSummaryВидеоOptimizing Spark Job PerformanceDIALOGUEImproving Apache Spark PerformanceЗадание
06Data Warehousing Fundamentals9 материалов

Unveiling Data Warehousing: From Concepts to Performance

Section IntroductionВидеоWhat Is a Data Warehouse (OLTP vs OLAP)ВидеоUnderstanding Data Warehouse, Data Lake, and LakehouseВидеоWhy Databricks Is Used for Analytics WorkloadsВидеоBasic Warehouse Concepts (Fact, Dimension, Star Schema – High Level)ВидеоHow Performance, Partitions & File Layout Matter in WarehousesВидеоSummaryВидеоUnderstanding Data Warehouse ConceptsDIALOGUECore Concepts of Data Warehousing and Modern AnalyticsЗадание
07Delta Lake & Data Management13 материалов

Mastering Delta Lake: From Data Foundations to Real-World Implementation

Section IntroductionВидеоProject Introduction: Business Use CaseВидеоWhat Is Delta Lake and Why It MattersВидеоUnderstanding ACID Transactions & Time TravelВидеоManaged vs External TablesВидеоDemonstration: Dataset PreparationВидеоDemonstration: End-to-End External Data Integration with S3 and Type of TablesВидеоDemonstration: Delta Lake ACID Transactions and Time TravelВидеоMedallion Architecture (Bronze–Silver–Gold)ВидеоDemonstration: End-to-End Medallion Architecture WalkthroughВидеоSummaryВидеоResolving Data Pipeline Issues with Delta LakeDIALOGUEDelta Lake Essentials and Data Pipelines in DatabricksЗадание
08Governance, BI & Pipelines10 материалов

Mastering Data Governance and Analytics in Databricks

Section IntroductionВидеоIntroduction to Unity Catalog (Why Governance Matters)ВидеоDemonstration: Data Governance in Databricks WorkspaceВидеоData Understanding & Simple Data Processing FlowВидеоDemonstration: Simple End-to-End Data Flow Using Delta Live TablesВидеоDemonstration: Understanding Delta Live Tables Pipeline SettingsВидеоDemonstration: Getting Insights with Genie & BI DashboardsВидеоSummaryВидеоUnderstanding Data Governance in DatabricksDIALOGUEModern Data Workflows in DatabricksЗадание
09Orchestration in Databricks7 материалов

Mastering Data Workflow Automation in Databricks

Section IntroductionВидеоDemonstration: Connecting GitHub with DatabricksВидеоIntroduction to Orchestration in DatabricksВидеоDemonstration: Creating & Scheduling Databricks JobsВидеоSummaryВидеоResolving a Job Scheduling Issue in DatabricksDIALOGUEOrchestration and Automation in DatabricksЗадание
10Capstone Project6 материалов

Building a Data Pipeline: From Ingestion to Visualization

Final Project WalkthroughВидеоDemonstration: Capstone Project Part-1ВидеоDemonstration: Capstone Project Part-2ВидеоDemonstration: Capstone Project Part-3ВидеоExploring Capstone Project ArchitectureDIALOGUETaxi Trip Data Analytics PipelineЗадание
11Conclusion3 материалов

Mastering the Fundamentals: A Data Engineer's Final Review

Course ConclusionВидеоPreparing for a Certification InterviewDIALOGUEThe Databricks Certified Data Engineer Associate - Practical Guide Final AssessmentЗадание