PyData Ann Arbor: Julia Silge | Text Mining Using Tidy Data Principles

2,181 views · Published 10 October 2018 · 1:00:56 · Indexed 22 September 2026

Channel: PyData · 2018 · Science & Technology

Watch on YouTube

PyData Ann Arbor Meetup - October 9, 2018
Sponsored by NumFOCUS and TD Ameritrade
https://www.meetup.com/PyData-Ann-Arbor/

Julia Silge | Text Mining Using Tidy Data Principles

Text data is increasingly important in many domains, and tidy data principles and tidy tools can make text mining easier and more effective. In this talk, Julia will demonstrate how we can manipulate, summarize, and visualize the characteristics of text using these methods and R packages from the tidy tool ecosystem. These tools are highly effective for many analytical questions and allow analysts to integrate natural language processing into effective workflows already in wide use. We will explore how to implement approaches such as sentiment analysis of texts, measuring tf-idf, and building text models.

-----

Julia Silge is a data scientist at Stack Overflow, with a PhD in astrophysics and an abiding love for Jane Austen. Julia worked in academia and ed tech before moving into data science and discovering R. She enjoys making beautiful charts, programming in R, text mining, and communicating about technical topics with diverse audiences. 00:00 Welcome!
00:10 Help us add time stamps or captions to this video! See the description for details.

Want to help add timestamps to our YouTube videos to help with discoverability? Find out more here: https://github.com/numfocus/YouTubeVideoTimestamps

More from this channel