Traseu de învățare
Inginer de date profesionist în Python
Creează-ți contul gratuit
Continuă cu GoogleArată mai multe opțiunisau
Îndrăgit de cursanți din mii de companii
Formare pentru o echipă?
Încearcă pentru afaceriDescrierea traseului
Inginer de date profesionist în Python
Cerințe prealabile
Inginer de dateCourse
Descoperă componentele cheie ale arhitecturii moderne de date, de la ingestie și servire la guvernanță și orchestrare.
Course
Course
Învață elementele esențiale ale VM-urilor, containerelor, Docker și Kubernetes. Înțelege diferențele pentru a începe!
Course
Acest curs prezintă dbt pentru modelarea datelor, transformări, testare și crearea documentației.
Course
Descoperă conceptele fundamentale ale programării orientate pe obiecte (OOP), construind clase și obiecte personalizate!
Course
Course
În această Introducere în DevOps, vei stăpâni bazele DevOps și vei învăța conceptele, instrumentele și tehnicile cheie pentru a îmbunătăți productivitatea.
Course
Stăpânește testarea Python: Învață metode, creează verificări și asigură cod fără erori cu pytest și unittest.
Project
bonusDebugging Code
Sharpen your debugging skills to enhance sales data accuracy.
Course
Descoperă Docker și importanța lui în setul de instrumente al profesionistului în date. Află despre containerele Docker, imaginile și altele.
Course
Stăpânește PySpark pentru a gestiona big data cu ușurință—învață să procesezi, interoghezi și optimizezi seturi de date uriașe pentru analize puternice!
Chapter
This chapter introduces the exciting world of Big Data, as well as the various concepts and different frameworks for processing Big Data. You will understand why Apache Spark is considered the best framework for BigData.
Chapter
The main abstraction Spark provides is a resilient distributed dataset (RDD), which is the fundamental and backbone data type of this engine. This chapter introduces RDDs and shows how RDDs can be created and executed using RDD Transformations and Actions.
Chapter
In this chapter, you'll learn about Spark SQL which is a Spark module for structured data processing. It provides a programming abstraction called DataFrames and can also act as a distributed SQL query engine. This chapter shows how Spark SQL allows you to use DataFrames in Python.
Project
Step into a data engineer's shoes and master data cleaning with PySpark on an e-commerce orders dataset!
Chapter
In this chapter, we learn how to download data files from web servers via the command line. In the process, we also learn about documentation manuals, option flags, and multi-file processing.
Chapter
In the last chapter, we bridge the connection between command line and other data science languages and learn how they can work together. Using Python as a case study, we learn to execute Python on the command line, to install dependencies using the package manager pip, and to build an entire model pipeline using the command line.
Course
Află diferența dintre batching și streaming, scalarea sistemelor de streaming și aplicațiile din lumea reală.
Course
Course
În acest curs, vei învăța bazele Kubernetes și vei implementa și orchestra containere folosind Manifests și instrucțiuni kubectl.
Resource
Understand how data engineering can impact your business.
finalizat
Obține diploma de absolvire
Adaugă această acreditare la profilul tău LinkedIn, CV sau rezumatDistribuie pe rețelele de socializare și în evaluarea ta de performanțăÎnscrie-te acum
Alătură-te celor peste 19 de milioane de cursanți și începe Inginer de date profesionist în Python astăzi!
Creează-ți contul gratuit
Continuă cu GoogleArată mai multe opțiunisau
Dezvoltați-vă abilitățile de gestionare a datelor cu DataCamp pentru mobil
Fă progrese din mers cu cursurile noastre mobile și provocările zilnice de programare de 5 minute.