Introduction to Text Analysis with Python

Text Analysis with Python

 

In Person

This two-class course is for people who have some experience coding in Python and would like to expand their capabilities to include the library Natural Language Toolkit (NLTK). Text analysis, a dynamic field within data science and linguistics, involves the systematic examination of textual data to uncover patterns, extract insights, and derive meaningful information. This process encompasses a range of techniques, from basic tasks like tokenization and stemming to more advanced methods such as sentiment analysis and named entity recognition. By leveraging tools like NLTK researchers can explore the structure and content of text, enabling a deeper understanding of language patterns and context. In this course, we are going to cover the basics of text analysis from pre-process corpora to simple analysis. If you are interested in developing further your text analysis skills, you can attend the Introduction to Topic Modelling with Bert course.

This is an intermediate-level course. You will need to already understand programming, preferably in Python. Previous knowledge of NLP is not required. It is also not required to have previous knowledge of Google Colab, although it is encouraged to set up a Google account prior to the workshops.

Those who have registered to take part will receive an email with full details on how to get ready for the course.

 

This course will be taught by Xan Cochran.

After taking part in this event, you may decide that you need some further help in applying what you have learnt to your research. If so, you can book a Data Surgery meeting with one of our training fellows.

More details about Data Surgeries. 

If you’re new to this training event format, or to CDCS training events in general, read more on what to expect from CDCS training. Here you will also find details of our cancellation and no-show policy, which applies to this event.

 

If you're interested in other training on text analysis, have a look at the following:

 

Return to the Training Homepage to see other available events.

Digital Scholarship Centre

Digital Scholarship Centre, 6th floor

Main Library 

University of Edinburgh 

Edinburgh EH8 9LJ

You might be interested in

an old map of Acotland with the text "Jennifer Smith & Brian Aitken, Project deep Dive"

Who Speaks Scots Where: What Crowdsourcing Reveals

Graphic for a workshop titled ‘Foundations of Webscraping.’ The background is a black-and-white photograph of students working together in a design studio with maps and models. A large teal ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner.

Collecting Data from the Web: Foundation of Webscraping

Graphic for a workshop titled ‘Text Classification in Practice: From Topic Models to Transformers.’ The background shows handwritten historical letters. A large green ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner.

Text Classification in Practice: From Topic Models to Transformers

Graphic for a workshop titled ‘Getting Started with Inferential Statistics.’ The background is a black-and-white photograph of people studying in a library with partitioned desks. A large teal ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner.

Getting Started with Inferential Statistics

Graphic for a workshop titled ‘Introduction to Geographical Data with QGIS.’ The background shows an old map of the world with detailed illustrations. A large teal ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner.

Intro to Geographical Data with QGIS

a yellow tinged photo of people entering a building, with the text "Brad Rittenhouse, Project Deep Dive"

Giving Humanists a Helping Hand in HPC

Graphic for a workshop titled ‘Working Collaboratively Through Version Control.’ The background is a black-and-white photograph of people weaving on large looms. A large magenta ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner.

Working Collaboratively through Version Control