Christian Bentz
Dados Biográficos
| ID | 201694 |
|---|---|
| NOME | Christian Bentz |
| PRENOMES | Christian |
| SOBRENOME | Bentz |
| ASSINATURA | BENTZ C |
| AFILIAÇÕES | University of Tübingen |
| ORCID | 0000-0001-6570-9326 |
| VERIFICADO | Sim |
| TOTAL DE OBRAS | 7 |
| TOTAL DE CITAÇÕES | 6 |
| TOTAL COMO AUTOR | 7 |
| TOTAL COMO EDITOR | 0 |
| PRIMEIRO ANO DE PUBLICAÇÃO | 2013 |
| ANO MAIS RECENTE DE PUBLICAÇÃO | 2023 |
| ÍNDICE H | 2 |
Complexity trade-offs and equi-complexity in natural languages
In linguistics, there is little consensus on how to define, measure, and compare complexity across languages. We propose to take the diversity of viewpoints as a given, and to capture the complexity of a language by a vector of measurements, rather than a single value. We then assess the statistical support for two controversial hypotheses: the trade-off hypothesis and the equi-complexity hypothesis. We furnish meta-analyses of 28 complexity metr…
Measuring language complexity
This special issue focuses on measuring language complexity. The contributions address methodological challenges, discuss implications for theoretical research, and use complexity measurements for testing theoretical claims. In this introductory article, we explain what knowledge can be gained from quantifying complexity. We then describe a workshop and a shared task which were our attempt to develop a systematic approach to the challenge of find…
Languages Through the Looking Glass of BPE Compression
Byte-pair encoding (BPE) is widely used in NLP for performing subword tokenization. It uncovers redundant patterns for compressing the data, and hence alleviates the sparsity problem in downstream applications. Subwords discovered during the first merge operations tend to have the most substantial impact on the compression of texts. However, the structural underpinnings of this effect have not been analyzed cross-linguistically. We conduct in-dep…
Meaning and Measures
Research on language complexity has been abundant and manifold in the past two decades. Within typology, it has to a very large extent been motivated by the question of whether all languages are equally complex, and if not, which language-external factors affect the distribution of complexity across languages. To address this and other questions, a plethora of different metrics and approaches has been put forward to measure the complexity of lang…
The evolution of language families is shaped by the environment beyond neutral drift
Modern human origins and dispersal
Languages with More Second Language Learners Tend to Lose Nominal Case
In this paper, we provide quantitative evidence showing that languages spoken by many second language speakers tend to have relatively small nominal case systems or no nominal case at all. In our sample, all languages with more than 50% second language speakers had no nominal case. The negative association between the number of second language speakers and nominal case complexity generalizes to different language areas and families. As there are …
Languages with More Second Language Learners Tend to Lose Nominal Case
In this paper, we provide quantitative evidence showing that languages spoken by many second language speakers tend to have relatively small nominal case systems or no nominal case at all. In our sample, all languages with more than 50% second language speakers had no nominal case. The negative association between the number of second language speakers and nominal case complexity generalizes to different language areas and families. As there are …
The evolution of language families is shaped by the environment beyond neutral drift
Modern human origins and dispersal
Meaning and Measures
Research on language complexity has been abundant and manifold in the past two decades. Within typology, it has to a very large extent been motivated by the question of whether all languages are equally complex, and if not, which language-external factors affect the distribution of complexity across languages. To address this and other questions, a plethora of different metrics and approaches has been put forward to measure the complexity of lang…
Complexity trade-offs and equi-complexity in natural languages
In linguistics, there is little consensus on how to define, measure, and compare complexity across languages. We propose to take the diversity of viewpoints as a given, and to capture the complexity of a language by a vector of measurements, rather than a single value. We then assess the statistical support for two controversial hypotheses: the trade-off hypothesis and the equi-complexity hypothesis. We furnish meta-analyses of 28 complexity metr…
Measuring language complexity
This special issue focuses on measuring language complexity. The contributions address methodological challenges, discuss implications for theoretical research, and use complexity measurements for testing theoretical claims. In this introductory article, we explain what knowledge can be gained from quantifying complexity. We then describe a workshop and a shared task which were our attempt to develop a systematic approach to the challenge of find…
Languages Through the Looking Glass of BPE Compression
Byte-pair encoding (BPE) is widely used in NLP for performing subword tokenization. It uncovers redundant patterns for compressing the data, and hence alleviates the sparsity problem in downstream applications. Subwords discovered during the first merge operations tend to have the most substantial impact on the compression of texts. However, the structural underpinnings of this effect have not been analyzed cross-linguistically. We conduct in-dep…
Computer Science (6 obras) · Language and cultural evolution (5 obras) · Sociology (4 obras) · Artificial Intelligence (3 obras) · Linguistics (3 obras) · Natural Language Processing Techniques (3 obras) · Psychology (3 obras) · Typology (3 obras) · Authorship Attribution and Profiling (2 obras) · Data science (2 obras)