ข้ามไปยังเนื้อหาหลัก

Courses

Categorical Data in the Tidyverse

พื้นฐานระดับทักษะ

อัปเดตแล้ว 01/2569

Get ready to categorize! In this course, you will work with non-numerical data, such as job titles or survey responses, using the Tidyverse landscape.

เริ่มเรียนหลักสูตรฟรี

RData Manipulation4 ชม.13 videos44 Exercises3,600 เอ็กซ์พี16,445คำแถลงแสดงความสำเร็จ

สร้างบัญชีฟรีของคุณ

หรือ

เมื่อดำเนินการต่อ คุณยอมรับข้อกำหนดการใช้งานของเรา นโยบายความเป็นส่วนตัวของเรา และยอมรับว่าข้อมูลของคุณจะถูกจัดเก็บไว้ในสหรัฐอเมริกา

เป็นที่ชื่นชอบของผู้เรียนในบริษัทหลายพันแห่ง

ฝึกอบรมบุคคลตั้งแต่ 2 คนขึ้นไป?

ลองใช้ DataCamp for Business

คำอธิบายรายวิชา

As a data scientist, you will often find yourself working with non-numerical data, such as job titles, survey responses, or demographic information. R has a special way of representing them, called factors, and this course will help you master working with them using the tidyverse package forcats. We’ll also work with other tidyverse packages, including ggplot2, dplyr, stringr, and tidyr and use real world datasets, such as the fivethirtyeight flight dataset and Kaggle’s State of Data Science and ML Survey. Following this course, you’ll be able to identify and manipulate factor variables, quickly and efficiently visualize your data, and effectively communicate your results. Get ready to categorize!

ข้อกำหนดเบื้องต้น

Reshaping Data with tidyr

1

Introduction to Factor Variables

In this chapter, you’ll learn all about factors. You’ll discover the difference between categorical and ordinal variables, how R represents them, and how to inspect them to find the number and names of the levels. Finally, you’ll find how forcats, a tidyverse package, can improve your plots by letting you quickly reorder variables by their frequency.

Introduction to qualitative variables

50 เอ็กซ์พี

Recognizing factor variables

100 เอ็กซ์พี

Qualitative variables in theory

50 เอ็กซ์พี

Understanding your qualitative variables

50 เอ็กซ์พี

Getting number of levels

100 เอ็กซ์พี

Examining number of levels

100 เอ็กซ์พี

Examining levels

100 เอ็กซ์พี

Making better plots

50 เอ็กซ์พี

Reordering a variable by its frequency

100 เอ็กซ์พี

Ordering one variable by another

100 เอ็กซ์พี

เริ่มบท

2

Manipulating Factor Variables

You’ll continue to dive into the forcats package, learning how to change the order and names of levels and even collapse them into one another.

Reordering factors

50 เอ็กซ์พี

Changing the order of factor levels

100 เอ็กซ์พี

Tricks of fct_relevel()

100 เอ็กซ์พี

Renaming factor levels

50 เอ็กซ์พี

Distinguishing between forcats functions

50 เอ็กซ์พี

Renaming a few levels

100 เอ็กซ์พี

When you have a typo

50 เอ็กซ์พี

Collapsing factor levels

50 เอ็กซ์พี

Manually collapsing levels

100 เอ็กซ์พี

Lumping variables by proportion

100 เอ็กซ์พี

Preserving the most common levels

100 เอ็กซ์พี

เริ่มบท

3

Creating Factor Variables

Having gotten a good grasp of forcats, you’ll expand out to the rest of the tidyverse, learning and reviewing functions from dplyr, tidyr, and stringr. You’ll refine graphs with ggplot2 by changing axes to percentage scales, editing the layout of the text, and more.

Examining common themed variables

50 เอ็กซ์พี

Grouping and reshaping similar columns

100 เอ็กซ์พี

Summarizing data

100 เอ็กซ์พี

Creating an initial plot

100 เอ็กซ์พี

Tricks of ggplot2

50 เอ็กซ์พี

Editing plot text

100 เอ็กซ์พี

Reordering graphs

100 เอ็กซ์พี

Changing and creating variables with case_when()

50 เอ็กซ์พี

case_when() with single variable

100 เอ็กซ์พี

case_when() from multiple columns

100 เอ็กซ์พี

เริ่มบท

4

Case Study on Flight Etiquette

In this final chapter, you’ll take all that you’ve learned and apply it in a case study. You’ll learn more about working with strings and summarizing data, then replicate a publication quality 538 plot.

Case study introduction

50 เอ็กซ์พี

Changing characters to factors

100 เอ็กซ์พี

Tidying data

100 เอ็กซ์พี

Data preparation and regex

50 เอ็กซ์พี

Cleaning up strings

100 เอ็กซ์พี

Dichotomizing variables

100 เอ็กซ์พี

Summarizing data

100 เอ็กซ์พี

Recreating the plot

50 เอ็กซ์พี

Creating an initial plot

100 เอ็กซ์พี

Fixing labels

100 เอ็กซ์พี

Flipping things around

100 เอ็กซ์พี

Finalizing the chart

100 เอ็กซ์พี

End of course recap

50 เอ็กซ์พี

เริ่มบท

Categorical Data in the Tidyverse

หลักสูตรเสร็จสมบูรณ์

ได้รับใบรับรองความสำเร็จ

เพิ่มข้อมูลรับรองนี้ลงในโปรไฟล์ LinkedIn, ประวัติย่อ หรือเรซูเม่ของคุณ
แชร์ลงในโซเชียลมีเดียและในรายงานประเมินผลการปฏิบัติงานของคุณลงทะเบียนเลย

สำหรับธุรกิจ

ฝึกอบรมบุคคลตั้งแต่ 2 คนขึ้นไป?

ให้ทีมของคุณเข้าถึงแพลตฟอร์ม DataCamp แบบเต็มรูปแบบ รวมถึงฟีเจอร์ทั้งหมด

ในแทร็กต่อไปนี้

Tidyverse Fundamentals in R

instructors

Emily Robinson

Senior Data Scientist, Game Data Pros

collaborators

Chester Ismay

Becca Robins

Courses resources

538 Flying Etiquette surveydatasets

Kaggle multiple choice responsesdatasets

เข้าร่วมกับ... 19 ล้านผู้เรียน และเริ่ม Categorical Data in the Tidyverse วันนี้เลย!

สร้างบัญชีฟรีของคุณ

หรือ

เมื่อดำเนินการต่อ คุณยอมรับข้อกำหนดการใช้งานของเรา นโยบายความเป็นส่วนตัวของเรา และยอมรับว่าข้อมูลของคุณจะถูกจัดเก็บไว้ในสหรัฐอเมริกา

พัฒนาทักษะด้านข้อมูลของคุณด้วย DataCamp for Mobile

พัฒนาทักษะได้ทุกที่ทุกเวลาด้วยคอร์สเรียนบนมือถือและแบบฝึกหัดเขียนโค้ดประจำวัน 5 นาทีของเรา