Exploratory Data Analysis with Textual Data in R / Quanteda

Offered By
Coursera Project Network
In this Guided Project, you will:

Learn how to import textual data, visualize textual data, stratify textual data by a third variable.

Clock2 hours
BeginnerBeginner
CloudNo download needed
VideoSplit-screen video
Comment DotsEnglish
LaptopDesktop only

In this 1-hour long project-based course, you will learn how to explore presidential concession speeches by US presidential candidates over time, looking specifically at speech length and top words and examining variation by Democrat and Republican candidates. You will learn how to import textual data stored in raw text files, turn these files into a corpus (a collection of textual documents) and tokenize the text all using the software package quanteda. You will also learn how to extract useful information from filenames and how to use this information to generate visualizations of textual data using the stringr and ggplot2 packages. Note: This course works best for learners who are based in the North America region. We’re currently working on providing the same experience in other regions.

Skills you will develop

Data AnalysisData Visualization (DataViz)R ProgrammingText Analysis

Learn step-by-step

In a video that plays in a split-screen with your work area, your instructor will walk you through these steps:

  1. You will learn how to import textual data stored in raw text files

  2. You will learn how to turn files into a corpus (a collection of textual documents)

  3. You will learn how to tokenize the text and turn text into a document feature matrix

  4. You will learn how to extract useful information from filenames

  5. You will learn how to generate visualizations of textual data

How Guided Projects work

Your workspace is a cloud desktop right in your browser, no download required

In a split-screen video, your instructor guides you step-by-step

Frequently asked questions

Frequently Asked Questions

More questions? Visit the Learner Help Center.