Text Analysis in R
Instructor: Alessandro Meneghini
Scheduled period: from 2 to 12 November 2026
Monday 2-11-2026, 13:00 - 17:00;
Wednesday 4-11-2026, 13:00 - 17:00;
Friday 6-11-2026, 13:00 - 17:00;
Tuesday 10-11-2026, 13:00 - 17:00;
Thursday 12-11-2026, 13:00 - 17:00;
Registration for the course will be open from 19 October at 9am to 23 October at 2pm at this link
An Open Badge will be issued by the University of Padova at the end of the course: https://bestr.it/badge/show/5893
Quantitative text analysis is a rapidly evolving field that involves the use of advanced computational tools to extract information from large amounts of textual data. The aim of the course is to provide participants with the skills needed to use advanced text analysis techniques, in order to extract insights from large textual datasets and to effectively interpret the results.
The course is based on R, a free and open-source programming language and development environment widely used in the scientific and academic community for data analysis. R offers a wide range of libraries and tools for text analysis.
The course will be structured to cover the main topics of quantitative text analysis, from corpus construction to the application of techniques such as topic modeling and sentiment analysis. By the end of the course, participants will be able to use R to analyze large amounts of text and to interpret the results effectively and consistently with the assumptions of this approach, acquiring a solid foundation for further study and application in the field of text analysis.
Course syllabus:
- Introduction to quantitative text analysis in R
- Corpus: Identification, construction, import and preprocessing
- Frequencies and Distributions
- Analysis techniques: Clustering, Topic Modeling, Sentiment Analysis
- Reading and interpreting results

