A new hybrid approach for automatic speech signal segmentation using silence signal detection, energy convex hull, and spectral variation

Xufang Zhao; D. O'Shaughnessy

doi:10.1109/CCECE.2008.4564512

Source

2008 Canadian Conference on Electrical and Computer Engineering > 145 - 000148

Abstract

This paper proposes a new approach for automatic syllable segmentation of Mandarin spontaneous speech. Automatic speech segmentation is important for continuous speech recognition because it reduces the search space effectively in automatic speech recognition. Moreover, the signal segmentation technique is useful in automatic speech marks and labels. However, for automatic speech recognition (ASR), it is difficult to segment the speech input reliably into useful sub-units because (1) syllable units can often be located roughly via intensity changes, but exact boundary positions are elusive in successive vowels, (2) energy changes in speech spectrum or amplitude help to estimate unit boundaries, but these cues are often unreliable due to co-articulation, and (3) finding boundaries for units bigger than phonemes combines the difficulties of detecting phoneme edges and of deciding which phonemes group to form the bigger units. In this paper, we present a hybrid segmentation method that utilizes silence detection, convex hull energy analysis, and spectral variation analysis. Furthermore, Hamming short-time sliding-windows were applied twice on audio signals to get more obvious convex hull valleys. Mandarin speech segmentation was used as a testing case, and the effectiveness of the proposed segmentation system was confirmed by the experimental results.

Identifiers

book ISSN :	0840-7789
book ISBN :	978-1-4244-1642-4
book e-ISBN :	978-1-4244-1643-1
DOI	10.1109/CCECE.2008.4564512

Keywords

speech synthesis natural language processing signal detection spectral analysis speech processing speech recognition Hamming short-time sliding-windows automatic speech signal segmentation silence signal detection automatic syllable segmentation Mandarin spontaneous speech automatic speech recognition successive vowels speech spectrum convex hull energy analysis spectral variation analysis Speech Equations Hidden Markov models Accuracy Detectors Mathematical model Acoustic signal analysis Acoustic signal processing

Additional information

Data set: ieee

Publisher

IEEE

INFONA - science communication portal

A new hybrid approach for automatic speech signal segmentation using silence signal detection, energy convex hull, and spectral variation

Source

Abstract

Identifiers

Authors

Xufang Zhao

O'Shaughnessy, D.

Keywords

Additional information

Publisher


Assign to other user
	×
Wrong email address

INFONA - science communication portal

A new hybrid approach for automatic speech signal segmentation using silence signal detection, energy convex hull, and spectral variation $("#expandableTitles").expandable();

Source

Abstract

Identifiers

Authors

User assignment

Assignment remove confirmation

You're going to remove this assignment. Are you sure?

Xufang Zhao

O'Shaughnessy, D.

Keywords

Additional information

Publisher

Share

Export to bibliography

Reporting an error / abuse

Sending the report failed

Accessibility options

A new hybrid approach for automatic speech signal segmentation using silence signal detection, energy convex hull, and spectral variation