Hans Van Halteren
Dados Biográficos
| ID | 739007 |
|---|---|
| NOME | Hans Van Halteren |
| PRENOMES | Hans |
| SOBRENOME | Van Halteren |
| ASSINATURA | VAN HALTEREN H |
| AFILIAÇÕES | Radboud University Nijmegen |
| ORCID | 0000-0001-8115-1799 |
| VERIFICADO | Sim |
| TOTAL DE OBRAS | 7 |
| TOTAL DE CITAÇÕES | 2 |
| TOTAL COMO AUTOR | 7 |
| TOTAL COMO EDITOR | 0 |
| PRIMEIRO ANO DE PUBLICAÇÃO | 1992 |
| ANO MAIS RECENTE DE PUBLICAÇÃO | 2023 |
| ÍNDICE H | 1 |
Why do we say them when we know it should be they ? Twitter as a resource for investigating nonstandard syntactic variation in The Netherlands
Two Twitter-based corpus studies are reported to account for the increasing preference in The Netherlands for the stigmatized subject use of the object pronoun hun ‘them.’ Twitter data were collected to obtain a sufficient number of hun -tokens, but also to investigate the validity of two hypotheses on the preference for hun , this is, that subject- hun is a contrast profiler which thrives in contexts of evaluation and qualification, and that sub…
Amsterdam, you're raining!” First-hand experience in tweets with spatio-temporal addressees
Improving Accuracy in Word Class Tagging through the Combination of Machine Learning Systems
We examine how differences in language models, learned by different data-driven systems performing the same NLP task, can be exploited to yield a higher accuracy than the best individual system. We do this by means of experiments involving the task of morphosyntactic word class tagging, on the basis of three different tagged corpora. Four well-known tagger generators (hidden Markov model, memory-based, transformation rules, and maximum entropy) a…
Title Index
The Linguistic Annotation of Corpora
The article discusses the role of linguistic annotation in corpus linguistics as opposed to annotation in natural language processing. In corpus linguistics, annotation is an integral part of the process of linguistic interpretation and description of the data. Tagging and parsing are discussed as the automatic counterparts of, respectively, the paradigmatic and the syntagmatic description of corpus data. The requirements for a corpus linguistic …
Linguistic Exploitation of Syntactic Databases
Computer-Grown Trees
Improving Accuracy in Word Class Tagging through the Combination of Machine Learning Systems
We examine how differences in language models, learned by different data-driven systems performing the same NLP task, can be exploited to yield a higher accuracy than the best individual system. We do this by means of experiments involving the task of morphosyntactic word class tagging, on the basis of three different tagged corpora. Four well-known tagger generators (hidden Markov model, memory-based, transformation rules, and maximum entropy) a…
Linguistic Exploitation of Syntactic Databases
Computer-Grown Trees
The Linguistic Annotation of Corpora
The article discusses the role of linguistic annotation in corpus linguistics as opposed to annotation in natural language processing. In corpus linguistics, annotation is an integral part of the process of linguistic interpretation and description of the data. Tagging and parsing are discussed as the automatic counterparts of, respectively, the paradigmatic and the syntagmatic description of corpus data. The requirements for a corpus linguistic …
Improving Accuracy in Word Class Tagging through the Combination of Machine Learning Systems
We examine how differences in language models, learned by different data-driven systems performing the same NLP task, can be exploited to yield a higher accuracy than the best individual system. We do this by means of experiments involving the task of morphosyntactic word class tagging, on the basis of three different tagged corpora. Four well-known tagger generators (hidden Markov model, memory-based, transformation rules, and maximum entropy) a…
Title Index
Amsterdam, you're raining!” First-hand experience in tweets with spatio-temporal addressees
Why do we say them when we know it should be they ? Twitter as a resource for investigating nonstandard syntactic variation in The Netherlands
Two Twitter-based corpus studies are reported to account for the increasing preference in The Netherlands for the stigmatized subject use of the object pronoun hun ‘them.’ Twitter data were collected to obtain a sufficient number of hun -tokens, but also to investigate the validity of two hypotheses on the preference for hun , this is, that subject- hun is a contrast profiler which thrives in contexts of evaluation and qualification, and that sub…
Computer Science (7 obras) · Linguistics (4 obras) · Philosophy (4 obras) · Artificial Intelligence (3 obras) · Natural language processing (3 obras) · Natural Language Processing Techniques (3 obras) · Philosophy (3 obras) · Corpus linguistics (2 obras) · Interpretation (philosophy (2 obras) · Language, Discourse, Communication Strategies (2 obras)