Exploratory Data Analysis with Textual Data in R / Quanteda

提供方
Coursera Project Network
在此指導項目中,您將:

Learn how to import textual data, visualize textual data, stratify textual data by a third variable.

Clock2 hours
Beginner初級
Cloud無需下載
Video分屏視頻
Comment Dots英語(English)
Laptop僅限桌面

In this 1-hour long project-based course, you will learn how to explore presidential concession speeches by US presidential candidates over time, looking specifically at speech length and top words and examining variation by Democrat and Republican candidates. You will learn how to import textual data stored in raw text files, turn these files into a corpus (a collection of textual documents) and tokenize the text all using the software package quanteda. You will also learn how to extract useful information from filenames and how to use this information to generate visualizations of textual data using the stringr and ggplot2 packages. Note: This course works best for learners who are based in the North America region. We’re currently working on providing the same experience in other regions.

您要培養的技能

Data AnalysisData Visualization (DataViz)R ProgrammingText Analysis

分步進行學習

在與您的工作區一起在分屏中播放的視頻中,您的授課教師將指導您完成每個步驟:

  1. You will learn how to import textual data stored in raw text files

  2. You will learn how to turn files into a corpus (a collection of textual documents)

  3. You will learn how to tokenize the text and turn text into a document feature matrix

  4. You will learn how to extract useful information from filenames

  5. You will learn how to generate visualizations of textual data

指導項目工作原理

您的工作空間就是瀏覽器中的雲桌面,無需下載

在分屏視頻中,您的授課教師會為您提供分步指導

常見問題

常見問題

還有其他問題嗎?請訪問 學生幫助中心