Quantitative Text Analysis and Measures of Readability in R

제공자:
Coursera Project Network
학습자는 이 안내 프로젝트에서 다음을 수행하게 됩니다.

Estimate the readability of a text document or corpus of documents.

Plot the variation in readability levels in a text corpus over time.

Clock1 hour
Beginner초급
Cloud다운로드 필요 없음
Video분할 화면 동영상
Comment Dots영어
Laptop데스크톱 전용

By the end of this project, you will be able to load textual data into R and turn it into a corpus object. You will also understand the concept of measures of readability in textual analysis. You will know how to estimate the level of readability of a text document or corpus of documents using a number of different readability metrics and how to plot the variation in readability levels in a text document corpus over time at the document and paragraph level. This project is aimed at beginners who have a basic familiarity with the statistical programming language R and the RStudio environment, or people with a small amount of experience who would like to learn how to measure the readability of textual data.

개발할 기술

  • Text Analysis
  • Data Wrangling
  • Data Visualization (DataViz)
  • Text Corpus
  • Readability

단계별 학습

작업 영역이 있는 분할 화면으로 재생되는 동영상에서 강사는 다음을 단계별로 안내합니다.

  1. Load textual data into R and turn it into a corpus object. You will also understand the concept of measures of readability in textual analysis.

  2. Estimate the level of readability of a text document or corpus of documents using a number of different readability metrics

  3. Prepare the textual data for plotting by extracting key information from text document filenames and combining these with readability data in a dataframe.

  4. Plot the variation in readability levels in a text document corpus over time.

  5. Reshape the data to paragraph level and plot the distribution of readability over time by paragraph.

안내형 프로젝트 진행 방식

작업 영역은 브라우저에 바로 로드되는 클라우드 데스크톱으로, 다운로드할 필요가 없습니다.

분할 화면 동영상에서 강사가 프로젝트를 단계별로 안내해 줍니다.

자주 묻는 질문

자주 묻는 질문

궁금한 점이 더 있으신가요? 학습자 도움말 센터를 방문해 보세요.