For realistic speech generation, variation in glottal waveform models has long been proposed. Due to simplicity and efficiency, the parametric models of the glottal flow are very popular in the field of speech generation. The proposed work presents a new approach to modeling the glottal flow. The current model is comprised of two piecewise differential equations that generate a glottal pulse. The first and second differential equations generate the opening and closing phases of the vocal folds, respectively while the closed phase is taken as zero. There are four parameters involved in the proposed model to bring variation in the shape of the glottal pulse. The current model is very flexible in designing a glottal pulse and is comparable with the famous Liljencrants-Fant model, Rosenberg model, and KLGLOTT88 model. This comparison supports its successful implementation as a voice source in speech synthesis which also leads to the validity of our differential equation-based glottal model.
Mushtaq Tahir, Murtaza Sadaf, Saleem Muhammad, Farhad Khurram, Ejaz Muhammad Faisal
. Parametric Modeling of Glottal Flow in Form of Piecewise Differential Equations[J]. Journal of Shanghai Jiaotong University(Science), 2026
, 31(4)
: 1048
-1056
.
DOI: 10.1007/s12204-025-2797-5
1. YANUSHEVSKAYA I, MURPHY A, GOBL C, et al. Global waveshape parameter rd in signaling focal prominence: Perceptual salience in the absence of f0 variation [J]. Frontiers in Communication, 2022, 7: 1026222.
2. FANT G, LILJENCRANTS J, LIN Q. A four-parameter model of glottal flow [J]. STL-QPSR, 1985, 26(4): 1-13.
3. LIA K, AKAGIA M, LIB Y, et al. Modeling and estimation of vocal tract and glottal source parameters using ARMAX-LF model [DB/OL]. (2024-10-07). https://arxiv.org/abs/2410.04704
4. FREIXES M, ARNELA M, SOCORÓ J C, et al. Glottal inverse filtering and vocal tract tuning for the numerical simulation of vowel/a/with different levels of vocal effort [C]//Interspeech 2024. Kos: ISCA, 2024: 3115-3119.
5. QURESHI T, SYED K. A one-mass physical model of the vocal folds with seesaw-like oscillations [J]. Archives of Acoustics, 2011, 36(1): 15-27.
6. THOMSON S L. Synthetic, self-oscillating vocal fold models for voice production research [J]. The Journal of the Acoustical Society of America, 2024, 156(2): 1283-1308.
7. GUASCH O. Control of chaotic vibrations in lumped mass models of the vocal folds [J]. INTER-NOISE and NOISE-CON Congress and Conference Proceedings, 2023, 267(1): 243-252.
8. JI M, LIU B, LOU Z, et al. Achievements and developments in mass models of vocal fold vibrations [J]. Journal of Shanghai Jiao Tong University (Science), 2023. https://doi.org/10.1007/s12204-023-2652-5
9. QURESHI T M, SYED K S. Fulcrum-point based self-oscillatory glottal model with numerical flow simulation[J]. International Journal of Acoustics & Vibration, 2018, 23(4): 516-528.
10. PERRINE B L, SCHERER R C. Using a vertical three-mass computational model of the vocal folds to match human phonation of three adult males [J]. The Journal of the Acoustical Society of America, 2023, 154(3): 1505-1525.
11. HAN Y X, ZHANG J X, WANG Y L. Dynamic behavior of a two-mass nonlinear fractional-order vibration system [J]. Frontiers in Physics, 2024, 12: 1452138.
12. ROTHENBERG M. Acoustic interaction between the glottal source and the vocal tract [J]. Vocal Fold Physiology, 1981, 1: 305-323.
13. PLUMPE M D, QUATIERI T F, REYNOLDS D A. Modeling of the glottal flow derivative waveform with application to speaker identification [J]. IEEE Transactions on Speech and Audio Processing, 1999, 7(5): 569-586.
14. LOBO A P. Glottal flow derivative modeling with the wavelet smoothed excitation [C]//2001 IEEE International Conference on Acoustics, Speech, and Signal Processing. Salt Lake City: IEEE, 2001: 861-864.
15. TITZE I R. Parameterization of the glottal area, glottal flow, and vocal fold contact area [J]. The Journal of the Acoustical Society of America, 1984, 75(2): 570-580.
16. PERROTIN O, FEUGÈRE L, D'ALESSANDRO C. Perceptual equivalence of the Liljencrants–Fant and linear-filter glottal flow models [J]. The Journal of the Acoustical Society of America, 2021, 150(2): 1273-1285.
17. ALKU P, STRIK H, VILKMAN E. Parabolic spectral parameter—a new method for quantification of the glottal flow [J]. Speech Communication, 1997, 22(1): 67-79.
18. ALKU P, VILKMAN E, LAUKKANEN A M. Parameterization of the voice source by combining spectral decay and amplitude features of the glottal flow [J]. Journal of Speech, Language, and Hearing Research, 1998, 41(5): 990-1002.
19. CORTÉS J P, ALZAMENDI G A, WEINSTEIN A J, et al. Kalman filter implementation of subglottal impedance-based inverse filtering to estimate glottal airflow during phonation [J]. Applied Sciences, 2022, 12(1): 401.
20. CABRAL J P, RENALS S, RICHMOND K, et al. Glottal spectral separation for parametric speech synthesis [C]//Interspeech 2008. Brisbane: ISCA, 2008: 1829-1832.
21. ROSENBERG A E. Effect of glottal pulse shape on the quality of natural vowels [J]. The Journal of the Acoustical Society of America, 1971, 49(2B): 583-590.
22. ROTHENBERG M, CARLSON R, GRANSTRÖM B, et al. A three-parameter voice source for speech synthesis [J]. Speech Communication, 1975, 2: 235-243.
23. FANT G. Glottal source and excitation analysis [J]. STL-QPSR, 1979, 20(1): 85-107.
24. HEDELIN P. A glottal LPC-vocode r[C]// IEEE International Conference on Acoustics, Speech, and Signal Processing. San Diego: IEEE, 1984: 21-24.
25. PRICE P J. Male and female voice source characteristics: Inverse filtering results [J]. Speech Communication, 1989, 8(3): 261-277.
26. VELDHUIS R. A computationally efficient alternative for the Liljencrants–Fant model and its perceptual evaluation [J]. The Journal of the Acoustical Society of America, 1998, 103(1): 566-571.
27. LI Y W, TAO J H, LIU B, et al. Comparison of glottal source parameter values in emotional vowels [C]//Interspeech 2020. Shanghai: ISCA, 2020: 4103-4107.
28. ZHANG Y, JIANG W L, SUN L N, et al. A deep-learning based generalized empirical flow model of glottal flow during normal phonation [J]. Journal of Biomechanical Engineering, 2022: 091001.
29. SCHULZE-FORSTER K, RICHARD G, KELLEY L, et al. Unsupervised music source separation using differentiable parametric source models [J]. IEEE/ACM Transactions on Audio, Speech, and Language Processing, 2023, 31: 1276-1289.
30. KLATT D H, KLATT L C. Analysis, synthesis, and perception of voice quality variations among female and male talkers [J]. The Journal of the Acoustical Society of America, 1990, 87(2): 820-857.
31. DOVAL B, D'ALESSANDRO C, HENRICH N. The spectrum of glottal flow models [J]. Acta Acustica United with Acustica, 2006, 92(6): 1026-1046.