Skip to main content

ETHNOS_APP

Home • Search • Journals • List 0

Modeling Lexical Tones for Speaker Discrimination

Bibliographic Data

ID6440693
AuthorsRicky Kw Chan (0000-0003-4977-8406, Speech, Language and Cognition Laboratory, School of English, The University of Hong Kong, Hong Kong, corresponding author), Bruce Xiao Wang (0000-0003-3564-2911, Hong Kong Polytechnic University)
Year2024
Volume68
Issue1
Pages229-243
Publication date2024-07-27
Peer ReviewedYes
Open AccessYes
TypeARTICLE
VenueLanguage and Speech (JOURNAL)
Journal identifiersISSN: 0023-8309 • E-ISSN: 1756-6053
PublisherSAGE Publishing (PUBLISHER • US)
DOI10.1177/00238309241261702
PMID39066631
OpenAlexW4401051652
LanguageEN
References cited30

Fundamental frequency (F0) has been widely studied and used in the context of speaker discrimination and forensic voice comparison casework, but most previous studies focused on long-term F0 statistics. Lexical tone, the linguistically structured and dynamic aspects of F0, has received much less research attention. A main methodological issue lies on how tonal F0 should be parameterized for the best speaker discrimination performance. This paper compares the speaker discriminatory performance of three approaches with lexical tone modeling: discrete cosine transform (DCT), polynomial curve fitting, and quantitative target approximation (qTA). Results show that using parameters based on DCT and polynomials led to similarly promising performance, whereas those based on qTA generally yielded relatively poor performance. Implications modeling surface tonal F0 and the underlying articulatory processes for speaker discrimination are discussed

Context (archaeology · Discrete cosine transform · Linguistics · Speaker recognition · Speech recognition · Tone (literature · Artificial Intelligence · Computer Science · Phonetics and Phonology Research · Speech and Audio Processing · Speech Recognition and Synthesis

  • Tone

    Open Access•Moira Yip•Tone•2002

  • Modern Cantonese Phonology

    Robert S Bauer, Paul K Benedict•Modern Cantonese Phonology•1997

  • An investigation of the effectiveness of a Swedish glide + vowel segment for speaker discrimination

    Erik Johannes Eriksson, Kirk P H Sullivan•International Journal of Speech…•2008

  • International Practices in Forensic Speaker Comparison

    Erica Gold, Peter French•International Journal of Speech…•2011

  • Speaker variability in the realisation of lexical tones

    Ricky K W Chan•International Journal of Speech…•2016

  • Does Lindley's LR estimation formula work for speech data? Investigation using long-term f0

    Yuko Kinoshita•International Journal of Speech…•2005

  • Background population

    Yuko Kinoshita, Shunichi Ishihara•International Journal of Speech…•2015

  • Exploring the Discriminatory Potential of F0 Distribution Parameters in Traditional Forensic Speaker Recognition

    Yuko Kinoshita, Shunichi Ishihara et al.•International Journal of Speech…•2009

  • Speaker-specific formant dynamics

    Kirsty Mcdougall•International Journal of Speech…•2004

  • Dynamic features of speech and the characterization of speakers

    Kirsty Mcdougall•International Journal of Speech…•2006

  • A New System of Cantonese Tones? Tone Perception and Production in Hong Kong South Asian Cantonese

    Open Access•Alan C L Yu, Crystal W T Lee et al.•Language and Speech•2022

Citation velocityhistorical
Highly citedNo

Tools

Open DOI
Ethnos_APP • Open Source Project • MIT License • Frontend v2.0.0 • Privacy and Cookies • API Documentation: api.ethnos.app/docs • API Source Code: GitHub • DOI: 10.5281/zenodo.17049435 • Frontend Source Code: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae