CLARIN Café | JOHD Special Collection on Language Dataset Reuse

Wednesday, 24 June 2026, 14:00 – 16:00 (CEST) – Online

The CLARIN Café dedicated to the Journal of Open Humanities Data Special Collection ‘Language Data Reuse: Opportunities, Challenges and Best Practices’ will bring together contributors and participants to discuss practical experiences, methodological questions, and infrastructural challenges surrounding the reuse of language data across the humanities.

The event aims to create a broader conversation on how different types of mono- and multilingual language resources, e.g. corpora, lexical resources, speech data, dialect data, historical datasets, and benchmarks for large language models, can be effectively reused, adapted, and extended for new research purposes.

Among the papers that will be presented, two papers by researchers from the Cnr-Istituto di Linguistica Computazionale “Antonio Zampolli” (CNR-ILC):

The programme includes also a “Guided discussion on experiences with data reuse, with focus on needs, bottlenecks and best practices” – Paola Marongiu (CNR-ILC) and Darja Fišer (CLARIN ERIC).

More info and registration: CLARIN Café | JOHD Special Collection on Language Dataset Reuse