ΑΙhub.org
 

History playground – finding patterns in historical newspapers


by
13 January 2020



share this:

Ever fancied finding out more about historical trends? Well, thanks to researchers at the University of Bristol, and their History Playground tool, anyone can analyse the content from a collection of historical British and American newspapers.

Macroscopic patterns of continuity and change over the course of centuries can be detected through the analysis of time series extracted from massive textual corpora. Similar data-driven approaches have already revolutionised the natural sciences. It is widely believed that there is similar potential for the humanities and social sciences. As such, new interactive tools are required to discover and extract macroscopic patterns from these vast quantities of data.

History Playground enables users to search for small sequences of words and retrieve their relative frequencies over the course of history. The tool makes use of scalable algorithms to first extract trends from textual corpora, before making them available for real-time search and discovery, presenting users with an interface to explore the data.

At present there are two large sets of text available:

Find out how to start using the History Playground by watching this short video:

Watch a further introduction to the project here:

History Playground uses the concept of n-grams, defined as short sequences of words. It is these n-grams that users search for when they use the tool. N-gram models are also widely used in the fields of natural language processing, probability, communication theory and data compression.

The team hope that in the long term, as more large textual datasets are released and additional feedback from the community helps to improve the Playground, they will be able to incorporate more varied and interesting corpora into the tool. In addition they are continuing to develop methods of analysis and additional views and visualisations. The tool also has the potential to incorporate text in languages other than English. For looking at more contemporary sources of data (for example, social media) the time resolution can be adjusted to study daily or even hourly changes.

This work is part of the ERC ThinkBIG project, Principal Investigator Nello Cristianini, University of Bristol.

Nello Cristianini is a Professor of Artificial Intelligence at the University of Bristol. His research interests include data science, artificial intelligence, machine learning, and applications to computational social sciences, digital humanities and news content analysis.

 

 

Read the full research articles on this topic:




Nello Cristianini is a Professor of Artificial Intelligence at the University of Bristol.
Nello Cristianini is a Professor of Artificial Intelligence at the University of Bristol.

            AUAI is supported by:



Subscribe to AIhub newsletter on substack



Related posts :

Everything, eco-where, AI at once?

Laura Martinez Agudelo builds on her research of visual representations of ecology and digitalisation to explore how "AI eco-imagery" is portrayed.

AI is making journalistic language more repetitive and predictable – and it’s a problem for all of us

  17 Jun 2026
What happens to language when a growing amount of text published in the press, online and on social media is written by machines?
monthly digest

AIhub monthly digest: June 2026 – biodiversity, resource allocation, and color metaphors

  16 Jun 2026
Welcome to our monthly digest, where you can catch up with AI research, events and news from the month past.

AAAI presidential panel – AI agents

  15 Jun 2026
Experts discuss AI agents, one of the topics covered in the AAAI Future of AI Research report.

Interview with AAAI Fellow Tanya Berger-Wolf: AI for ecology, biodiversity, and conservation

  11 Jun 2026
Find out about Tanya work on a foundation model for biology and the insights that this can provide.

Statistical or embodied? Comparing people and LLMs in their processing of color metaphors: an interview with Douglas Guilbeault

  09 Jun 2026
We learn what implications color metaphors and synaesthesia have for human and AI cognition.

The Good Robot podcast: the battle over data centres with Tara Merk

  08 Jun 2026
Eleanor Drage speaks with Tara Merk about how community-owned data centers could transform digital ownership and challenge the dominance of Big Tech.

Congratulations to the #AAMAS2026 best paper award winners

  05 Jun 2026
Find out who won in the categories of best paper, best student paper, and best blue sky paper.



AUAI is supported by:







Subscribe to AIhub newsletter on substack




 















©2026.05 - Association for the Understanding of Artificial Intelligence