Assessing phoneme distribution for speech modeling

Jesús A. Parra; Carlos Calvache; Matías Zañartu

doi:10.1117/12.2670042

6 March 2023 Assessing phoneme distribution for speech modeling

Jesús A. Parra, Carlos Calvache, Matías Zañartu

Proceedings Volume 12567, 18th International Symposium on Medical Information Processing and Analysis; 1256716 (2023) https://doi.org/10.1117/12.2670042
Event: 18th International Symposium on Medical Information Processing and Analysis, 2022, Valparaíso, Chile

Abstract

Phonetically balanced texts are used to study different voice and speech characteristics. In the context of clinical work and research, these texts provide a standard for quantifying perceptual, acoustic, or aerodynamic assessments. Recent modeling efforts are being devoted to describing long-term speech behaviors based on a collection of sustained phonemes. However, comprehensive descriptions of phoneme distributions representative of connected speech are not readily available. Thus, the present study introduces a method to estimate phoneme distributions using text data mining, as an alternative to existing power law methods. The procedure used for the decomposition of texts into phonemes, the estimation of the phonetic distributions and the comparisons between different texts, conversational speech, and standard reading passages are discussed. The results are presented using histograms and R-squared determination coefficients for the case of the English language, although the approach can be easily applied for other languages. A discussion of the proposed method, results, and limitations is presented.

Citation Download Citation

Jesús A. Parra, Carlos Calvache, and Matías Zañartu "Assessing phoneme distribution for speech modeling", Proc. SPIE 12567, 18th International Symposium on Medical Information Processing and Analysis, 1256716 (6 March 2023); https://doi.org/10.1117/12.2670042

ACCESS THE FULL ARTICLE

INSTITUTIONAL
Select your institution to access the SPIE Digital Library.

SELECT YOUR INSTITUTION

PERSONAL
Sign in with your SPIE account to access your personal subscriptions or to use specific features such as save to my library, sign up for alerts, save searches, etc.

PERSONAL SIGN IN

No SPIE Account? Create one

PURCHASE THIS CONTENT

SUBSCRIBE TO DIGITAL LIBRARY

50 downloads per 1-year subscription

Members: $195

Non-members: $335 ADD TO CART

25 downloads per 1 - year subscription

Members: $145

Non-members: $250 ADD TO CART

PURCHASE SINGLE ARTICLE

Includes PDF, HTML & Video, when available

Members: $17.00

Non-members: $21.00 ADD TO CART

PROCEEDINGS
7 PAGES

DOWNLOAD PAPER SAVE TO MY LIBRARY

GET CITATION

RIGHTS & PERMISSIONS

Get copyright permission Get copyright permission on Copyright Marketplace

KEYWORDS

Histograms

Modeling

Acoustics

Diseases and disorders

RELATED CONTENT

Modelling of dielectric elastomer loudspeakers including dissipative effects
Proceedings of SPIE (April 09 2013)

Feature analysis acoustic signals for fiber optic sensing based NDE...
Proceedings of SPIE (June 13 2023)

Fuzzy logic, edge enabled underwater video surveillance through partially wireless...
Proceedings of SPIE (October 17 2023)

A novel deployable structure devised from a Kresling-Scissor interface
Proceedings of SPIE (April 18 2023)

mTITAN: multi-domain tactical intelligent teaming and autonomous navigation
Proceedings of SPIE (June 12 2023)

Path planning of UAV pesticide spraying in terraced fields based...
Proceedings of SPIE (October 19 2023)

Acoustic metamaterial containing an array of Helmholtz resonators coupled with...
Proceedings of SPIE (April 22 2020)

Subscribe to Digital Library

Receive Erratum Email Alert

Keywords/Phrases

Search In:

Publication Years