Skip to main content

ETHNOS_APP

Home • Search • Journals • List 0

Os desafios para disponibilização e compartilhamento de dados linguísticos da Amostra Base Varsul

Bibliographic Data

ID3821513
AuthorsIsabel De Oliveira E Silva Monguilhott (0000-0003-1705-4390, Universidade Federal de Santa Catarina), Izete Lehmkuhl Coelho (0000-0001-6865-6004, Universidade Federal de Santa Catarina), Cláudia Regina Brescancini (0000-0003-4950-494X, Pontifícia Universidade Católica do Rio Grande do Sul)
Year2025
Volume6
Issue4
Publication date2025-10-02
Peer ReviewedYes
Open AccessYes
TypeARTICLE
VenueCadernos de Linguística e Teoria da Literatura (Universidade Federal de Minas Gerais) (JOURNAL)
Journal identifiersISSN: 2675-4916 • E-ISSN: 2675-4916
PublisherAssociacao Brasileira de Linguistica (PUBLISHER)
DOI10.25189/2675-4916.2025.v6.n4.id813
OpenAlexW4414759799
LanguagePT

This article aims to present the challenges of making available and sharing linguistic data from the VARSUL Base Sample. The VARSUL project was set up in the 1980s to study the Portuguese spoken in the southern region of Brazil, including data from the capital cities and the most historically and socio-culturally important urban centres. It currently brings together linguists from four universities in the southern region: the Federal University of Rio Grande do Sul (UFRGS), the Pontifical Catholic University of Rio Grande do Sul (PUCRS), the Federal University of Santa Catarina (UFSC) and the Federal Technological University of Paraná (UTFPR). To make up the Basic Sample of the VARSUL database, 288 personal experience interviews were carried out between 1989 and 1996 - 96 per state and 24 in each of the 12 cities selected – taking into account the ethnic groups that make up the regions. In addition to ethnicity, the sample was stratified by gender (male and female), two age groups (25 to 49; over 50) and three levels of schooling (4 to 5 years; 8 to 9 years and 10 to 11 years). All 288 interviews were transcribed and stored at the VARSUL project branch offices and have served as rich material for describing and analysing the varieties of Southern Brazilian Portuguese. The database has complementary samples to expand the Base Sample, as well as samples from urban and non-urban neighbourhoods in Florianópolis-SC. In order to make the linguistic data of the Base Sample publicly accessible, in accordance with the guidelines of Open Science, we are currently in the process of de-identifying and anonymizing the audio recordins and transcripts of the interviews, in order to guarantee the confidentiality of the participants, in accordance with ethical and legal requirements. Also on the current agenda of the VARSUL project is the implementation of the project ‘Study of Linguistic Change in Real Time: expansion of the VARSUL project speech database’, whose main goal is to expand the base sample through panel studies and trend studies, with the aim of studying change in real time

Ethnic group · Federal state · Portuguese · Natural Language Processing Techniques

Citation velocityhistorical
Highly citedNo

Tools

Open DOIOpen Access
Ethnos_APP • Open Source Project • MIT License • Frontend v2.0.0 • Privacy and Cookies • API Documentation: api.ethnos.app/docs • API Source Code: GitHub • DOI: 10.5281/zenodo.17049435 • Frontend Source Code: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae