Extraction of vocal-tract system characteristics from speech signals

B. Yegnanarayana, Raymond N.J. Veldhuis

    Research output: Contribution to journalArticleAcademicpeer-review

    116 Citations (Scopus)
    774 Downloads (Pure)

    Abstract

    We propose methods to track natural variations in the characteristics of the vocal-tract system from speech signals. We are especially interested in the cases where these characteristics vary over time, as happens in dynamic sounds such as consonant-vowel transitions. We show that the selection of appropriate analysis segments is crucial in these methods, and we propose a selection based on estimated instants of significant excitation. These instants are obtained by a method based on the average group-delay property of minimum-phase signals. In voiced speech, they correspond to the instants of glottal closure. The vocal-tract system is characterized by its formant parameters, which are extracted from the analysis segments. Because the segments are always at the same relative position in each pitch period, in voiced speech the extracted formants are consistent across successive pitch periods. We demonstrate the results of the analysis for several difficult cases of speech signals
    Original languageUndefined
    Pages (from-to)313-327
    Number of pages15
    JournalIEEE transactions on speech and audio processing
    Volume6
    Issue number4
    DOIs
    Publication statusPublished - Jul 1998

    Keywords

    • glottal closure
    • IR-55938
    • pitch periods
    • average group-delay property
    • speech signals
    • vocal-tract system characteristics
    • voiced speech
    • EWI-15209
    • Analysis segments
    • dynamic sounds
    • natural variations
    • extracted formants
    • estimated significant excitation instants
    • formant parameters
    • consonant-vowel transitions
    • minimum-phase signals

    Cite this