Wednesday, 24 June 2026, 14:00 – 16:00 (CEST) – Online
The CLARIN Café dedicated to the Journal of Open Humanities Data Special Collection ‘Language Data Reuse: Opportunities, Challenges and Best Practices’ will bring together contributors and participants to discuss practical experiences, methodological questions, and infrastructural challenges surrounding the reuse of language data across the humanities.
The event aims to create a broader conversation on how different types of mono- and multilingual language resources, e.g. corpora, lexical resources, speech data, dialect data, historical datasets, and benchmarks for large language models, can be effectively reused, adapted, and extended for new research purposes.
Among the papers that will be presented, two papers by researchers from the Cnr-Istituto di Linguistica Computazionale “Antonio Zampolli” (CNR-ILC):
- Making sense of Italian legacy computational lexical resources: A sense annotation experience (Valeria Quochi, Francesca Frontini and Monica Monachini)
- Reviving Legacy WordNet-like Resources: MariTerm and ItalWordNet Renewal through Mutual Expansion and Plug-in Links (Lucia Galiero, Angelo Mario Del Grosso and Monica Monachini)
The programme includes also a “Guided discussion on experiences with data reuse, with focus on needs, bottlenecks and best practices” – Paola Marongiu (CNR-ILC) and Darja Fišer (CLARIN ERIC).
