Emmanuel Dupoux
Datos Biográficos
| ID | 757183 |
|---|---|
| NOMBRE | Emmanuel Dupoux |
| NOMBRES | Emmanuel |
| APELLIDO | Dupoux |
| FIRMA | DUPOUX E |
| AFILIACIONES | Centre National de la Recherche Scientifique |
| ORCID | 0000-0002-7814-2952 |
| VERIFICADO | Sí |
| TOTAL DE OBRAS | 62 |
| TOTAL DE CITAS | 13 |
| TOTAL COMO AUTOR | 61 |
| TOTAL COMO EDITOR | 1 |
| PRIMER AÑO DE PUBLICACIÓN | 1987 |
| AÑO MÁS RECIENTE DE PUBLICACIÓN | 2025 |
| ÍNDICE H | 2 |
Simulating Early Phonetic and Word Learning Without Linguistic Categories
Before they even talk, infants become sensitive to the speech sounds of their native language and recognize the auditory form of an increasing number of words. Traditionally, these early perceptual changes are attributed to an emerging knowledge of linguistic categories such as phonemes or words. However, there is growing skepticism surrounding this interpretation due to limited evidence of category knowledge in infants. Previous modeling work ha…
Modeling early phonetic acquisition from child-centered audio data
Modeling the initial state of early phonetic learning in infants
What are the necessary conditions to acquire language? Do infants rely on simple statistical mechanisms, or do they come pre-wired with innate capabilities allowing them to learn their native language(s)? Previous modeling studies have shown that unsupervised learning algorithms could reproduce some aspects of infant phonetic learning. Despite these successes, algorithms still fail to reproduce the learning trajectories observed in infants. Here,…
Modeling early phonetic acquisition from child-centered audio data
Infants learn their native language(s) at an amazing speed. Before they even talk, their perception adapts to the language(s) they hear. However, the mechanisms responsible for this perceptual attunement and the circumstances in which it takes place remain unclear. This paper presents the first attempt to study perceptual attunement using ecological child-centered audio data. We show that a simple prediction algorithm exhibits perceptual attuneme…
Realistic and broad-scope learning simulations
There is a current 'theory crisis' in language acquisition research, resulting from fragmentation both at the level of the approaches and the linguistic level studied. We identify a need for integrative approaches that go beyond these limitations, and propose to analyse the strengths and weaknesses of current theoretical approaches of language acquisition. In particular, we advocate that language learning simulations, if they integrate realistic …
How much does prosody help word segmentation? A simulation study on infant-directed speech
Statistical learning models of early phonetic acquisition struggle with child-centered audio data
Infants learn their native language(s) at an amazing speed. Before they even talk, their perception adapts to the language(s) they hear. However, the mechanisms responsible for this perceptual attunement still remain unclear. A long tradition in linguistics points to the importance of specialized language mechanisms that would allow us to quickly and effortlessly learn from the language(s) we are exposed to. However, the currently dominant explan…
Can statistical learning bootstrap early language acquisition? A modeling investigation
Before they even produce their first word, infants start developing a language-specific perception, recognize the auditory form of frequent words, and develop a rudimentary knowledge of grammatical categories. A major question in language development is: what mechanisms are responsible for the effortless learning infants demonstrate? In-laboratory experiments have shown that young infants are exquisitely sensitive to fine-grained statistical regu…
The effect of different information sources on prosodic boundary perception
This study aims to quantify the effect of several information sources: acoustic, higher-level linguistic, and knowledge of the prosodic system of the language, on the perception of prosodic boundaries. An experiment with native and non-native participants investigating the identification of prosodic boundaries in Japanese was conducted. It revealed that non-native speakers as well as native speakers with access only to acoustic information can re…
SCALa
Theories and data on language acquisition suggest a range of cues are used, ranging from information on structure found in the linguistic signal itself, to information gleaned from the environmental context or through social interaction. We propose a blueprint for computational models of the early language learner (SCALa, for Socio-Computational Architecture of Language Acquisition) that makes explicit the connection between the kinds of informat…
Does Infant‐Directed Speech Help Phonetic Learning? A Machine Learning Investigation
A prominent hypothesis holds that by speaking to infants in infant-directed speech (IDS) as opposed to adult-directed speech (ADS), parents help them learn phonetic categories. Specifically, two characteristics of IDS have been claimed to facilitate learning: hyperarticulation, which makes the categories more separable, and variability, which makes the generalization more robust. Here, we test the separability and robustness of vowel category lea…
Reverse engineering language acquisition with child-centered long-form recordings
Language use in everyday life can be studied using lightweight, wearable recorders that collect long-form recordings - that is, audio (including speech) over whole days. The hardware and software underlying this technique is increasingly accessible and inexpensive, and these data are revolutionizing the language acquisition field. We first place this technique into the broader context of the current ways of studying both the input being received …
Do Infants Really Learn Phonetic Categories
Early changes in infants' ability to perceive native and nonnative speech sound contrasts are typically attributed to their developing knowledge of phonetic categories. We critically examine this hypothesis and argue that there is little direct evidence of category knowledge in infancy. We then propose an alternative account in which infants' perception changes because they are learning a perceptual space that is appropriate to represent speech, …
Reverse Engineering Language Acquisition with Child-Centered Long-Form Recordings
Language use in everyday life can be studied using lightweight, wearable recorders that collect long-form recordings—that is, audio (including speech) over whole days. The hardware and software underlying this technique are increasingly accessible and inexpensive, and these data are revolutionizing the language acquisition field. We first place this technique into the broader context of the current ways of studying both the input being received b…
Segmentability Differences Between Child-Directed and Adult-Directed Speech
Previous computational modeling suggests it is much easier to segment words from child-directed speech (CDS) than adult-directed speech (ADS). However, this conclusion is based on data collected in the laboratory, with CDS from play sessions and ADS between a parent and an experimenter, which may not be representative of ecologically collected CDS and ADS. Fully naturalistic ADS and CDS collected with a nonintrusive recording device as the child …
WordSeg
A basic task in first language acquisition likely involves discovering the boundaries between words or morphemes in input where these basic units are not overtly segmented. A number of unsupervised learning algorithms have been proposed in the last 20 years for these purposes, some of which have been implemented computationally, but whose results remain difficult to compare across papers. We created a tool that is open source, enables reproducibl…
Early phonetic learning without phonetic categories -- Insights from large-scale simulations on realistic input
Before they even speak, infants become attuned to the sounds of the language(s) they hear, processing native phonetic contrasts more easily than non-native ones. For example, between 6-8 months and 10-12 months, infants learning American English get better at distinguishing English [ɹ] and [l], as in ‘rock’ vs ‘lock’, relative to infants learning Japanese. Influential accounts of this early phonetic learning phenomenon initially proposed that inf…
Are Words Easier to Learn From Infant‐ Than Adult‐Directed Speech? A Quantitative Corpus‐Based Investigation
We investigate whether infant-directed speech (IDS) could facilitate word form learning when compared to adult-directed speech (ADS). To study this, we examine the distribution of word forms at two levels, acoustic and phonological, using a large database of spontaneous speech in Japanese. At the acoustic level we show that, as has been documented before for phonemes, the realizations of words are more variable and less discriminable in IDS than …
Relating unsupervised word segmentation to reported vocabulary acquisition
A range of computational approaches have been used to model the discovery of word forms from continuous speech by infants. Typically, these algorithms are evaluated with respect to the ideal ’gold standard’ word segmentation and lexicon. These metrics assess how well an algorithm matches the adult state, but may not reflect the intermediate states of the child’s lexical development. We set up a new evaluation method based on the correlation betwe…
ASR Systems as Models of Phonetic Category Perception in Adults
We test the potential of standard Automatic Speech Recognition (ASR) systems trained on large corpora of continuous speech as quantitative models of human speech processing. In human adults, speech perception is attuned to efficiently process native speech sounds, at the expense of difficulties in pro- cessing non-native sounds. We use ABX-discriminability measures to test whether ASR models can account for the patterns of confusion between speec…
The more, the better? Behavioral and neural correlates of frequent and infrequent vowel exposure
A central assumption in the perceptual attunement literature holds that exposure to a speech sound contrast leads to improvement in native speech sound processing. However, whether the amount of exposure matters for this process has not been put to a direct test. We elucidated indicators of frequency-dependent perceptual attunement by comparing 5-8-month-old Dutch infants' discrimination of tokens containing a highly frequent [hɪt-he:t] and a hig…
Child-Directed Speech Is Infrequent in a Forager-Farmer Population
This article provides an estimation of how frequently, and from whom, children aged 0-11 years (Ns between 9 and 24) receive one-on-one verbal input among Tsimane forager-horticulturalists of lowland Bolivia. Analyses of systematic daytime behavioral observations reveal < 1 min per daylight hour is spent talking to children younger than 4 years of age, which is 4 times less than estimates for others present at the same time and place. Adults prov…
Frequency & Compositionality in Emergent Communication
In natural languages, frequency and compositionality exhibit an inverse relationship: the most frequent words often resist regular patterns, developing idiosyncratic forms. This phenomenon, exemplified by irregular verbs, raises a compelling question: do artificial communication systems follow similar principles? Through systematic experiments with neural network agents in a referential game setting, and by manipulating input frequency through Zi…
LongTail-Swap
Children learn how to speak with a low amount of data and can be taught new words on a few-shot basis, which makes them particularly data-efficient learners. The BabyLM challenge aims at exploring language model (LM) training in the low data regime but uses metrics that concentrate on the head of the word distribution. Here, we introduce LongTail-Swap (LT-Swap), a benchmark that focuses on the tail of the distribution, i.e., measures the ability …
Coping With Linguistic Diversity
Because multilingual environments are the norm rather than the excep tion, infants must have the capacity very early in life to distinguish one language from another. Without such a capacity, infants might acquire linguistic systems that amalgamate properties of different languages. The ensuing confusion would be overpowering. Fortunately, this never arises, despite the intuitive fears of monolingual parents. Infants raised in multilingual socie…
Perception of predictable stress
Reverse Engineering Language Acquisition with Child-Centered Long-Form Recordings
Language use in everyday life can be studied using lightweight, wearable recorders that collect long-form recordings—that is, audio (including speech) over whole days. The hardware and software underlying this technique are increasingly accessible and inexpensive, and these data are revolutionizing the language acquisition field. We first place this technique into the broader context of the current ways of studying both the input being received b…
De la psychologie à la science cognitive
De la psychologie à la science cognitive
Monitoring the lexicon with normal and compressed speech
Bootstrapping lexical acquisition
Les As. tentent de determiner comment les nourrissons peuvent se construire un lexique en ne percevant qu'un signal de parole continu et en ne pouvant pas recourir a l'identification lexicale, comme les adultes. Ils proposent un modele de perception de la parole base sur l'idee que la prosodie est utilisee par les nourrissons (et par les adultes) afin d'effectuer une segmentation du flux de parole. L'acquisition lexicale serait ainsi amorcee sur …
A Destressing “Deafness” in French
Where Is the Length Effect? A Cross-Linguistic Study of Speech Production
Epenthetic vowels in Japanese
In four cross-linguistic experiments comparing French and Japanese hearers, we found that the phonotactic properties of Japanese (very reduced set of syllable types) induce Japanese listeners to perceive ``illusory'' vowels inside consonant clusters in VCCV stimuli. In Experiments 1 and 2, we used a continuum of stimuli ranging from no vowel (e.g. ebzo) to a full vowel between the consonants (e.g. ebuzo). Japanese, but not French participants, re…
A robust method to study stress “deafness”
Previous research by Dupoux et al. [J. Memory Lang. 36, 406–421 (1997)] has shown that French participants, as opposed to Spanish participants, have difficulties in distinguishing nonwords that differ only in the location of stress. Contrary to Spanish, French does not have contrastive stress, and French participants are “deaf” to stress contrasts. The experimental paradigm used by Dupoux et al. (speeded ABX) yielded significant group differences…
Language, Brain, and Cognitive Development
Interdisciplinary essays on central issues in cognitive science. In the early 1960s, the bold project of the emerging field of cognition was to put the human mind under the scrutiny of rational inquiry, through the conjoined efforts of philosophy, linguistics, computer science, psychology, and neuroscience. Forty years later, cognitive science is a flourishing academic field. The contributions to this collection, written in honor of Jacques Mehle…
Testing Infants' Discrimination With the Orientation Latency Procedure
A new discrimination procedure based on the measurement of visual orientation latency to speech stimuli is introduced. Each participant listens to a series of short familiarization test trials. In each trial, 5 to 7 centrally‐presented familiarization stimuli are followed by laterally‐presented test stimuli. Infants were found to orient faster to different‐category than to same‐category test stimuli. This result was found despite a high degree of…
An Influence of Syntactic and Semantic Variables on Word Form Retrieval
We report the case of DPI, an aphasic patient who shows a phonological impairment in production that spares certain syntactic and semantic categories. On a picture naming task, he produces mostly phono-logical paraphasias, and the probability of producing a correct response depends on the frequency and length of the target word. This deficit occurs in the presence of spared ability to find the grammatical gender of the items that he cannot name, …
Partial Awareness Creates the “Illusion” of Subliminal Semantic Priming
We argue that the lack of consensus regarding the existence of subliminal semantic processing arises from not taking into account the fact that linguistic stimuli are represented across several processing levels (features, letters, word form) that can independently reach or not reach awareness. Using masked words, we constructed conditions in which participants were aware of some letters or fragments of a word, while remaining unaware of the whol…
Subliminal Speech Priming
We present a novel subliminal priming technique that operates in the auditory modality. Masking is achieved by hiding a spoken word within a stream of time-compressed speechlike sounds with similar spectral characteristics. Participants were unable to consciously identify the hidden words, yet reliable repetition priming was found. This effect was unaffected by a change in the speaker's voice and remained restricted to lexical processing. The res…
Misperception in sentences but not in words
We report two case studies of aphasic patients with a working-memory impairment due to reduced storage in the phonological buffer. The two patients display excellent performance in phonological discrimination tasks as long as the tasks do not involve a memory load. We then show that their performance drops when they have to maintain fine-grained phonological information for sentence comprehension: They are impaired at mispronunciation detection a…
Breaking the mirror
In this paper, we study the link between the processing systems that sustain speech perception and production in a patient (F.A.) with conduction aphasia. Her pattern of performance in repetition task - quantitative but also qualitative striking difference in errors with pseudowords versus words - cannot be properly accounted for either by a perception deficit or by a production deficit. We discuss this finding according to theoretical models of …
The native language of social cognition
What leads humans to divide the social world into groups, preferring their own group and disfavoring others? Experiments with infants and young children suggest these tendencies are based on predispositions that emerge early in life and depend, in part, on natural language. Young infants prefer to look at a person who previously spoke their native language. Older infants preferentially accept toys from native-language speakers, and preschool chil…
Consciousness and Cognition
What were the circumstances that led to the development of our cognitive abilities from a primitive hominid to an essentially modern human? The answer to this question is of profound importance to understanding our present nature. Since the steep path of our cognitive development is the attribute that most distinguishes humans from other mammals, this is also a quest to determine human origins. This collection of outstanding scientific problems a…
Apprentissage « bottom-up » des phonèmes
Nous présentons une étude computationnelle de l'hypothèse selon laquelle l'information distri-butionnelle est suffisante pour acquérir les règles allophoniques (et ainsi les phonèmes) de façon bottom-up. L'hypothèse a été testée en utilisant une mesure de la théorie de l'information qui compare les distributions. La phase de test a été conduite sur plusieurs corpus de langues artificielles et sur deux corpus de langues naturelles (constitués de t…
How “semantic” is response priming restricted to practiced items? A reply to Abrams & Grinspan (2007)
Persistent stress ‘deafness’
Accent Over Race
Episodic accessibility and morphological processing
Language‐specific stress perception by 9‐month‐old French and Spanish infants
During the first year of life, infants begin to have difficulties perceiving non-native vowel and consonant contrasts, thus adapting their perception to the phonetic categories of the target language. In this paper, we examine the perception of a non-segmental feature, i.e. stress. Previous research with adults has shown that speakers of French (a language with fixed stress) have great difficulties in perceiving stress contrasts (Dupoux, Pallier,…
How rich is consciousness? The partial awareness hypothesis
The development of a phonological illusion
In adults, native language phonology has strong perceptual effects. Previous work has shown that Japanese speakers, unlike French speakers, break up illegal sequences of consonants with illusory vowels: they report hearing abna as abuna. To study the development of phonological grammar, we compared Japanese and French infants in a discrimination task. In Experiment 1, we observed that 14-month-old Japanese infants, in contrast to French infants, …
Holographic String Encoding
In this article, we apply a special case of holographic representations to letter position coding. We translate different well-known schemes into this format, which uses distributed representations and supports constituent structure. We show that in addition to these brain-like characteristics, performances on a standard benchmark of behavioral effects are improved in the holographic format relative to the standard localist one. This notably occu…
Psychology (52 obras) · Computer Science (38 obras) · Linguistics (38 obras) · Language Development and Disorders (34 obras) · Cognitive psychology (32 obras) · Phonetics and Phonology Research (30 obras) · Perception (21 obras) · Artificial Intelligence (20 obras) · Speech recognition (20 obras) · Natural language processing (18 obras)