Skip to main content

Course

ETL and ELT in Python

Intermediate4 hr

Learn to build effective, performant, and reliable data pipelines using Extract, Transform, and Load principles.

Python4 hr14 videos53 Exercises4,450 XP38,298Statement of accomplishment

Create Your Free Account

Continue with Google
or
By continuing, you accept our Terms of Use, our Privacy Policy and that your data is stored in the USA.

Loved by learners at thousands of companies

Training a Team?

Try for Business

Course Description

Empowering Analytics with Data Pipelines

Data pipelines are at the foundation of every strong data platform. Building these pipelines is an essential skill for data engineers, who provide incredible value to a business ready to step into a data-driven future. This introductory course will help you hone the skills to build effective, performant, and reliable data pipelines.

Building and Maintaining ETL Solutions

Throughout this course, you’ll dive into the complete process of building a data pipeline. You’ll grow skills leveraging Python libraries such as pandas and json to extract data from structured and unstructured sources before it’s transformed and persisted for downstream use. Along the way, you’ll develop confidence tools and techniques such as architecture diagrams, unit-tests, and monitoring that will help to set your data pipelines out from the rest. As you progress, you’ll put your new-found skills to the test with hands-on exercises.

Supercharge Data Workflows

After completing this course, you’ll be ready to design, develop and use data pipelines to supercharge your data workflow in your job, new career, or personal project.

Feels like what you want to learn?

Start Course for Free

What you'll learn

  • Assess data integrity and pipeline performance using logging, validation checkpoints, and automated unit or end-to-end tests
  • Differentiate ETL and ELT architectures in terms of process sequence, tooling, and appropriate storage targets
  • Evaluate deployment and orchestration options that schedule, monitor, and retry pipelines in production environments
  • Identify the essential stages and components of Python-based data pipelines, including data sources, transformations, and destinations
  • Recognize pandas and SQL techniques for extracting, transforming, and loading both tabular and non-tabular datasets

Prerequisites

Curriculum

Course outline

1

Introduction to Data Pipelines

Get ready to discover how data is collected, processed, and moved using data pipelines. You will explore the qualities of the best data pipelines, and prepare to design and build your own.
Start Chapter

ETL and ELT in Python

Course
Complete

Earn Statement of Accomplishment

Enroll Now

Grow your data skills with DataCamp for Mobile

Make progress on the go with our mobile courses and daily 5-minute coding challenges.