A value-sensitive metadata schema for interpreting corpora
Implementation on the Unified Interpreting Corpus (Unic) platform
Dados Bibliográficos
| ID | 6152306 |
|---|---|
| Autores | Nannan Liu (0000-0001-9235-2633, University of Bologna), Mariachiara Russo (0000-0002-0904-2771, University of Bologna) |
| Ano | 2025 |
| Volume | 27 |
| Fascículo | 2 |
| Páginas | 157-196 |
| Data de publicação | 2025-08-29 |
| Peer Reviewed | Sim |
| Open Access | Sim |
| Tipo | ARTICLE |
| Periódico | Interpreting International Journal of Research and Practice in Interpreting (JOURNAL) |
| Identificadores do periódico | ISSN: 1384-6647 • E-ISSN: 1569-982X |
| Editora | John Benjamins Publishing Company (PUBLISHER • NL) |
| DOI | 10.1075/intp.00123.liu |
| OpenAlex | W4413826845 |
| Idioma | IT |
| Referências citadas | 43 |
Interpreting corpora serve as the descriptive foundation of research and the ‘ground truth’ against which machine interpreting technologies are evaluated. However, access to corpora remains a critical bottleneck in interpreting studies due to data collection and processing challenges and the absence of interpreting- and translation-specific corpus publication venues. In this article, we present two technical infrastructures that facilitate corpus access: a metadata schema which standardises corpus description and the Unified Interpreting Corpus (UNIC) platform for data and metadata search and publication. Guided by the internationally established FAIR (findability, accessibility, interoperability and reusability) and CARE (collective benefit, authority to control, responsibility and ethics) principles for scientific data management and stewardship, we designed the infrastructures based on a review of 125 spoken and signed language interpreting corpora, relevant international standards and community knowledge and also by using open-source technologies. Feedback obtained from interpreting students, researchers and interpreters demonstrates greater perceived usefulness of and satisfaction with UNIC compared to general-purpose search portals. Overall, we illustrate a value- and consensus-driven path towards optimising the use of interpreting corpora and the careful curation of new ones, which avoids the duplication of effort, helps to chart research directions and fosters co-design with communities
Information retrieval · Metadata · Schema (genetic algorithms · World Wide Web · Computer Science · Interpreting and Communication in Healthcare · Natural Language Processing Techniques · Translation Studies and Practices
The Use of Databases in Cross-Linguistic Studies
Value Sensitive Design
The Care Principles for Indigenous Data Governance
The Fair Guiding Principles for scientific data management and stewardship
Is machine interpreting interpreting
Iris
Examining simultaneous interpreting norms and strategies in a South African legislative context
Corpus-based translation studies
Du texte aux ressources multimodales
A Descriptive Study of Norms in Interpreting
Speaking in the first-person singular or plural
A corpus for signed language interpreting research
Corpus-based Interpreting Studies as an Offshoot of Corpus-based Translation Studies
Gender and gender agreement in bilingual native and non-native grammars
Seven Dimensions of Portability for Language Documentation and Description
| Velocidade de citação | historical |
|---|---|
| Altamente citado | Não |