Finding Patterns Across Data

patterns

 

IN PERSON 

This course focuses on how to find patterns of similarity and clusters in data. The course will cover both the techniques on how to find the underlying patterns in data using R/RStudio, alongside ways of visualising and presenting these patterns. 

Each session in the workshop begins with a presentation introducing theoretical principles. This is followed by a guided practical exercise where these principles can be tested and applied, as well as developed upon. 

The first class focuses on finding clusters and groups in data and will cover the following: 

  • How to extract Principal Components using PCA 

  • Clustering with K-Means 

  • Hierarchical Clustering using the UPGMA algorithm 

The second class focuses on measuring diversity within datasets and will cover the following: 

  • How to quantify diversity 

  • How to compute Jaccard Similarity Coefficients 

  • How to compute the Morisita-Horn overlap coefficient 

  • Visualising pairwise similarities 

The course will develop a foundational understanding of cluster analysis, similarity and diversity indices, and the functionality of R/RStudio through interactive participation in each session. By fostering familiarity with the basic principles and software, it will allow attendees to confidently apply these, or indeed seek out new applications, for their own and future research. 

This is an intermediate-level course. Intermediate sessions explore specific aspects of the method (libraries, tools etc.) Students must have a basic background in R. This includes, at least the basic data types in R, how to install and load packages, and, more generally, of how the R Studio interface works.  

Those who have registered to take part will receive an email with full details on how to get ready for the workshop. 

After taking part in this event, you may decide that you need some further help in applying what you have learnt to your research. If so, you can book a Data Surgery meeting with one of our training fellows. 

More details about Data Surgeries. 

If you’re new to this training event format, or to CDCS training events in general, read more on what to expect from CDCS training. Here you will also find details of our cancellation and no-show policy, which applies to this event. 

 

If you're interested in other training on data analysis, statistics, and machine learning have a look at the following: 

Return to the Training Homepage to see other available events.

Digital Scholarship Centre

Digital Scholarship Centre, 6th floor

Main Library 

University of Edinburgh 

Edinburgh EH8 9LJ

You might be interested in

Graphic for a workshop titled ‘Introduction to Geographical Data with QGIS.’ The background shows an old map of the world with detailed illustrations. A large teal ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner.

Intro to Geographical Data with QGIS

an old map of Acotland with the text "Jennifer Smith & Brian Aitken, Project deep Dive"

Who Speaks Scots Where: What Crowdsourcing Reveals

Graphic for a workshop titled ‘Getting Started with Inferential Statistics.’ The background is a black-and-white photograph of people studying in a library with partitioned desks. A large teal ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner.

Getting Started with Inferential Statistics

Graphic for an event titled ‘BYOD Festival.’ The background is a black-and-white photograph of people sitting around a table, drinking tea and playing cards. A large magenta ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner.

Bring Your Own Data (BYOD) Fest

black and white photograph of a person drinking tea out of a flaks on top of a hill.

CDCS December Fika

Graphic for a workshop titled ‘Using API for Research.’ The background is a black-and-white photograph of people working with printing equipment and patterned sheets. A large magenta ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner.

Using API for Research

Graphic for a workshop titled ‘Data Viscualisation’ The background is a collage of historical printed text with an overlaid image of a wolf. A large green ampersand featuring an illustration of Ada Lovelace is placed on the left. The logo of the Centre for Data, Culture & Society (DCS) appears in the top right corner

Digital Method of the Month: Data Visualisation