The Common Language Resources and Tools Infrastructure (CLARIN) is a major pan-European research infrastructure which is building services to underpin the interoperability, re-use and sustainability of language data and software for research across the humanities and social sciences. One of the key aims of CLARIN is to help researchers who develop useful datasets and software to be able to make them more easily available for others to find, access, use, combine with other resources, and to help to keep them available in the long term.

CLARIN is built on expert centres, and national funding initiatives to support and sustain these centres. The CLARIN European Research Infrastructure Consortium (ERIC) is the body which co-ordinates these activities and provides the central services to glue them together, and is based at the University of Utrecht in the Netherlands. European countries opt in to build their national CLARIN infrastructure and to contribute to the costs of running the ERIC. So far, nineteen counties have joined as full members, with two observers. Much more information is available at the CLARIN website.

CLARIN could potentially help us to share our resources more widely more effectively. The UK joined the CLARIN ERIC as an Observer in 2015, renewed for a second three-year period in 2018.  Please get in touch if you'd like to join in - email Martin Wynne.

CLARIN is a European network. These pages give you information and links to data, applications, services and tools in the UK, but the real strength of CLARIN comes from sharing across borders. Please visit the CLARIN European Research Infrastructure Consortium gateway to services for access to many more digital language resources.

Data & Tools

Screenshot of the EIST online interface

ÈIST Gaelic speech recognition tool

Interactional Variation Online logo, without text

IVO Corpus

Interactional Variation Online Corpus

Latest News

A student presenting her poster at the CLARIN annual conference in 2025

PhD student session

Library stock image

Organising and Discovering Texts in the Oxford Text Archive with Subcollections

Kate Hawkes

Events

clarin2026 square 1080 x 1080 px

CLARIN Annual Conference 2026

Logo for CLARIN-DARIAH UK

CLARIN and DARIAH UK Annual Event

Find language resources via the CLARIN Virtual Language Observatory

The CLARIN Virtual Language Observatory pulls together records from all the CLARIN centres, and many other language resource repositories.

Use the search bar below to enter a search term and start searching through hundreds of thousands of language resources via a search term, or just hit 'Search the CLARIN VLO' button to browse everything and use facets to narrow down to your area of interest or discover new resources.

 

Search the CLARIN VLO

 

Find out more about the CLARIN VLO and other CLARIN services here.

To play the video, you need to accept cookies for videos.

CLARIN-UK Consortium Members

Photo British Library, St Pancras. CLARIN

British Library

Centre for Corpus Research, University of Birmingham

Centre for Corpus Research

University of Birmingham
University of Leeds Language Centre

Centre for Translation Studies

University of Leeds
Corpws Cenedlaethol Cymraeg Cyfoes logo

CorCenCC

The Welsh National Corpus team at Cardiff University
Image of an old printed book

CTTR University of Wolverhampton

Wolverhampton University
University of Oxford logo

Faculty of Linguistics, Philology and Phonetics

University of Oxford
King's Digital Lab abstract logo

King's College London

King's Digital Lab and Department of Digital Humanities
Lancaster University

Lancaster University

CASS and UCREL
Photograph of the Edinburgh skyline

Language Technology Group

University of Edinburgh
Natural Language Processing Group, University of Sheffield

NLPG

Natural Language Processing Group, University of Sheffield
Image from research project in cognitive linguistics, University of Glasgow

School of Critical Studies

University of Glasgow
Coventry University Logo

School of Humanities

Coventry University