Massimo Poesio
Biographic Data
| ID | 6213010 |
|---|---|
| NAME | Massimo Poesio |
| GIVEN NAMES | Massimo |
| FAMILY NAME | Poesio |
| SIGNATURE | POESIO M |
| AFFILIATIONS | Queen Mary University of London |
| ORCID | 0000-0001-8469-2072 |
| VERIFIED | Yes |
| TOTAL WORKS | 10 |
| TOTAL CITATIONS | 59 |
| AUTHOR COUNT | 10 |
| EDITOR COUNT | 0 |
| FIRST PUBLICATION YEAR | 2000 |
| LATEST PUBLICATION YEAR | 2024 |
| H-INDEX | 3 |
Polysemy—Evidence from Linguistics, Behavioral Science, and Contextualized Language Models
Polysemy is the type of lexical ambiguity where a word has multiple distinct but related interpretations. In the past decade, it has been the subject of a great many studies across multiple disciplines including linguistics, psychology, neuroscience, and computational linguistics, which have made it increasingly clear that the complexity of polysemy precludes simple, universal answers, especially concerning the representation and processing of po…
Anaphoric reference to mereological entities
Corpus evidence suggests that in contexts in which the presence of multiple antecedents might favor plural reference, the disadvantage observed for singular reference may disappear if the potential antecedents are combined in a group-like plural entity. We examined the relative salience of antecedents in conditions where the context either made a group interpretation available (i.e., mereological entity) (e.g., The engineer hooked up the engine t…
Computational Models of Anaphora
Interpreting anaphoric references is a fundamental aspect of our language competence that has long attracted the attention of computational linguists. The appearance of ever-larger anaphorically annotated data sets covering more and more anaphoric phenomena in ever-greater detail has spurred the development of increasingly more sophisticated computational models; as a result, the most recent state-of-the-art neural models are able to achieve impr…
Pandora’s Box Opened
Janet Hitzeman
Evaluating Centering for Information Ordering Using Corpora
In this article we discuss several metrics of coherence defined using centering theory and investigate the usefulness of such metrics for information ordering in automatic text generation. We estimate empirically which is the most promising metric and how useful this metric is using a general methodology applied on several corpora. Our main result is that the simplest metric (which relies exclusively on NOCB transitions) sets a robust baseline th…
Inter-Coder Agreement for Computational Linguistics
This article is a survey of methods for measuring agreement among corpus annotators. It exposes the mathematics and underlying assumptions of agreement coefficients, covering Krippendorff's alpha as well as Scott's pi and Cohen's kappa; discusses the use of coefficients in several annotation tasks; and argues that weighted, alpha-like coefficients, traditionally less used than kappa-like measures in computational linguistics, may be more appropri…
Underspecification and Anaphora: Theoretical Issues and Preliminary Evidence
Much experimental work in psycholinguistics suggests that fully specified syntactic and semantic interpretations are obtained incrementally. The finding that interpretation takes place incrementally is very robust and underlies our own view of sentence processing as well; however, most of this work tends to test very simple interpretive judgments using materials that have clean-cut interpretations, which makes the earlier-expressed view more ques…
Centering: A Parametric Theory and Its Instantiations
Centering theory is the best-known framework for theorizing about local coherence and salience; however, its claims are articulated in terms of notions which are only partially specified, such as “utterance,” “realization,” or “ranking.” A great deal of research has attempted to arrive at more detailed specifications of these parameters of the theory; as a result, the claims of centering can be instantiated in many different ways. We investigated…
An Empirically Based System for Processing Definite Descriptions
We present an implemented system for processing definite descriptions in arbitrary domains. The design of the system is based on the results of a corpus analysis previously reported, which highlighted the prevalence of discourse-new descriptions in newspaper corpora. The annotated corpus was used to extensively evaluate the proposed techniques for matching definite descriptions with their antecedents, discourse segmentation, recognizing discourse…
Inter-Coder Agreement for Computational Linguistics
This article is a survey of methods for measuring agreement among corpus annotators. It exposes the mathematics and underlying assumptions of agreement coefficients, covering Krippendorff's alpha as well as Scott's pi and Cohen's kappa; discusses the use of coefficients in several annotation tasks; and argues that weighted, alpha-like coefficients, traditionally less used than kappa-like measures in computational linguistics, may be more appropri…
Centering: A Parametric Theory and Its Instantiations
Centering theory is the best-known framework for theorizing about local coherence and salience; however, its claims are articulated in terms of notions which are only partially specified, such as “utterance,” “realization,” or “ranking.” A great deal of research has attempted to arrive at more detailed specifications of these parameters of the theory; as a result, the claims of centering can be instantiated in many different ways. We investigated…
Underspecification and Anaphora: Theoretical Issues and Preliminary Evidence
Much experimental work in psycholinguistics suggests that fully specified syntactic and semantic interpretations are obtained incrementally. The finding that interpretation takes place incrementally is very robust and underlies our own view of sentence processing as well; however, most of this work tends to test very simple interpretive judgments using materials that have clean-cut interpretations, which makes the earlier-expressed view more ques…
An Empirically Based System for Processing Definite Descriptions
We present an implemented system for processing definite descriptions in arbitrary domains. The design of the system is based on the results of a corpus analysis previously reported, which highlighted the prevalence of discourse-new descriptions in newspaper corpora. The annotated corpus was used to extensively evaluate the proposed techniques for matching definite descriptions with their antecedents, discourse segmentation, recognizing discourse…
Computational Models of Anaphora
Interpreting anaphoric references is a fundamental aspect of our language competence that has long attracted the attention of computational linguists. The appearance of ever-larger anaphorically annotated data sets covering more and more anaphoric phenomena in ever-greater detail has spurred the development of increasingly more sophisticated computational models; as a result, the most recent state-of-the-art neural models are able to achieve impr…
An Empirically Based System for Processing Definite Descriptions
We present an implemented system for processing definite descriptions in arbitrary domains. The design of the system is based on the results of a corpus analysis previously reported, which highlighted the prevalence of discourse-new descriptions in newspaper corpora. The annotated corpus was used to extensively evaluate the proposed techniques for matching definite descriptions with their antecedents, discourse segmentation, recognizing discourse…
Centering: A Parametric Theory and Its Instantiations
Centering theory is the best-known framework for theorizing about local coherence and salience; however, its claims are articulated in terms of notions which are only partially specified, such as “utterance,” “realization,” or “ranking.” A great deal of research has attempted to arrive at more detailed specifications of these parameters of the theory; as a result, the claims of centering can be instantiated in many different ways. We investigated…
Underspecification and Anaphora: Theoretical Issues and Preliminary Evidence
Much experimental work in psycholinguistics suggests that fully specified syntactic and semantic interpretations are obtained incrementally. The finding that interpretation takes place incrementally is very robust and underlies our own view of sentence processing as well; however, most of this work tends to test very simple interpretive judgments using materials that have clean-cut interpretations, which makes the earlier-expressed view more ques…
Evaluating Centering for Information Ordering Using Corpora
In this article we discuss several metrics of coherence defined using centering theory and investigate the usefulness of such metrics for information ordering in automatic text generation. We estimate empirically which is the most promising metric and how useful this metric is using a general methodology applied on several corpora. Our main result is that the simplest metric (which relies exclusively on NOCB transitions) sets a robust baseline th…
Inter-Coder Agreement for Computational Linguistics
This article is a survey of methods for measuring agreement among corpus annotators. It exposes the mathematics and underlying assumptions of agreement coefficients, covering Krippendorff's alpha as well as Scott's pi and Cohen's kappa; discusses the use of coefficients in several annotation tasks; and argues that weighted, alpha-like coefficients, traditionally less used than kappa-like measures in computational linguistics, may be more appropri…
Janet Hitzeman
Pandora’s Box Opened
Computational Models of Anaphora
Interpreting anaphoric references is a fundamental aspect of our language competence that has long attracted the attention of computational linguists. The appearance of ever-larger anaphorically annotated data sets covering more and more anaphoric phenomena in ever-greater detail has spurred the development of increasingly more sophisticated computational models; as a result, the most recent state-of-the-art neural models are able to achieve impr…
Anaphoric reference to mereological entities
Corpus evidence suggests that in contexts in which the presence of multiple antecedents might favor plural reference, the disadvantage observed for singular reference may disappear if the potential antecedents are combined in a group-like plural entity. We examined the relative salience of antecedents in conditions where the context either made a group interpretation available (i.e., mereological entity) (e.g., The engineer hooked up the engine t…
Polysemy—Evidence from Linguistics, Behavioral Science, and Contextualized Language Models
Polysemy is the type of lexical ambiguity where a word has multiple distinct but related interpretations. In the past decade, it has been the subject of a great many studies across multiple disciplines including linguistics, psychology, neuroscience, and computational linguistics, which have made it increasingly clear that the complexity of polysemy precludes simple, universal answers, especially concerning the representation and processing of po…
Computer Science (10 works) · Artificial Intelligence (7 works) · Linguistics (7 works) · Natural language processing (7 works) · Natural Language Processing Techniques (7 works) · Philosophy (5 works) · Interpretation (philosophy (4 works) · Philosophy (4 works) · Psychology (4 works) · Speech and dialogue systems (4 works)