Search results

Items from 1 to 6 out of 6 results

chapter

Modified MPE/MMI in a transducer-based framework

G. Heigold, R. Schluter, H. Ney

2009 IEEE International Conference on Acoustics, Speech and Signal Processing > 3749 - 3752

ICASSP 2009 - 2009 IEEE International Conference on Acoustics, Speech and Signal Processing

In this paper we show how common training criteria like for example MPE or MMI can be extended to incorporate a margin term. In addition, a transducer-based training implementation is presented, which covers a large variety of discriminative training criteria for ASR, including the standard MMI, MPE, and MCE criteria, as well as the modifications to these criteria presented here. The modified criteria...

chapter

Multiple time resolution analysis of speech signal using MCE training with application to speech recognition

S. Dimopoulos, A. Potamianos, E.-F. Lussier, Chin-Hui Lee

2009 IEEE International Conference on Acoustics, Speech and Signal Processing > 3801 - 3804

ICASSP 2009 - 2009 IEEE International Conference on Acoustics, Speech and Signal Processing

In this paper, we propose two methods of multiple time-resolution analysis of speech and their application to automatic speech recognition (ASR). Constant frame-rate multi-scale analysis is proposed based on a box of multi-scale features. Then a variable rate analysis is proposed based on the selection of the optimal temporal resolution on the fly by a properly trained non-linear classifier unit....

article

Discriminative learning in sequential pattern recognition

Xiaodong He, Li Deng, Wu Chou

IEEE Signal Processing Magazine > 2008 > 25 > 5 > 14 - 36

In this article, we studied the objective functions of MMI, MCE, and MPE/MWE for discriminative learning in sequential pattern recognition. We presented an approach that unifies the objective functions of MMI, MCE, and MPE/MWE in a common rational-function form of (25). The exact structure of the rational-function form for each discriminative criterion was derived and studied. While the rational-function...

chapter

An enhanced minimum classification error learning framework for balancing insertion, deletion and substitution errors

Yuan Fu Liao, Jia Jang Tu, Sen Chia Chang, Chin Hui Lee

2007 IEEE Workshop on Automatic Speech Recognition&Understanding (ASRU) > 587 - 590

2007 IEEE Workshop on Automatic Speech Recognition and Understanding

In continuous speech recognition substitution, insertion and deletion errors usually not only vary in numbers but also have different degrees of impact on optimizing a set of acoustic models. To balance their contributions to the overall error, an enhanced minimum classification error (E-MCE) learning framework is developed. The basic idea is to partition acoustic model optimization into three subtasks,...

chapter

Comparison of Large Margin Training to Other Discriminative Methods for Phonetic Recognition by Hidden Markov Models

Fei Sha, Lawrence K. Saul

2007 IEEE International Conference on Acoustics, Speech and Signal Processing - ICASSP '7 > 4 > IV-313 - IV-316

2007 IEEE International Conference on Acoustics, Speech, and Signal Processing

In this paper we compare three frameworks for discriminative training of continuous-density hidden Markov models (CD-HMMs). Specifically, we compare two popular frameworks, based on conditional maximum likelihood (CML) and minimum classification error (MCE), to a new framework based on margin maximization. Unlike CML and MCE, our formulation of large margin training explicitly penalizes incorrect...

chapter

A Novel Learning Method for Hidden Markov Models in Speech and Audio Processing

Xiaodong He, Li Deng, Wu Chou

2006 IEEE Workshop on Multimedia Signal Processing > 80 - 85

2006 IEEE Workshop on Multimedia Signal Processing

In recent years, various discriminative learning techniques for HMMs have consistently yielded significant benefits in speech recognition. In this paper, we present a novel optimization technique using the minimum classification error (MCE) criterion to optimize the HMM parameters. Unlike maximum mutual information training where an extended Baum-Welch (EBW) algorithm exists to optimize its objective...

Filter options

Data set:
ieee
Keywords:
SPEECH RECOGNITION

Publication date

Set your own date range

Publication type

book (5)
article (1)

Keywords

HIDDEN MARKOV MODELS (4)
TRAINING (4)
LEARNING (ARTIFICIAL INTELLIGENCE) (3)
MMI (3)
AUTOMATIC SPEECH RECOGNITION (2)
DISCRIMINATIVE LEARNING (2)
LARGE MARGIN (2)
MINIMUM CLASSIFICATION ERROR (2)
MPE (2)
OPTIMIZATION (2)
PATTERN CLASSIFICATION (2)
SPEECH (2)
SPEECH PROCESSING (2)
ACCURACY (1)
ACOUSTICS (1)
ASR (1)
AUDIO PROCESSING (1)
AUDIO SIGNAL PROCESSING (1)
CONDITIONAL RANDOM FIELDS (1)
CONTINUOUS MANDARIN DIGIT RECOGNITION (1)
CONTINUOUS SPEECH RECOGNITION SUBSTITUTION (1)
DISCRIMINATIVE LEARNING TECHNIQUE (1)
DISCRIMINATIVE TRAINING (1)
DISTANCE MEASUREMENT (1)
EBW ALGORITHM (1)
ENHANCED MINIMUM CLASSIFICATION ERROR LEARNING (1)
ERROR STATISTICS (1)
EXTENDED BAUM-WELCH (1)
EXTENDED BAUM-WELCH ALGORITHM (1)
FASTENERS (1)
FINITE STATE MACHINES (1)
FINITE STATE TRANSDUCER-BASED TRAINING FRAMEWORK (1)
FRAME-WISE CLASSIFICATION TASK (1)
GROWTH TRANSFORMATION (1)
HIDDEN MARKOV MODEL (1)
HMM (1)
JOINTS (1)
MACHINE LEARNING (1)
MANDARIN DIGIT RECOGNITION (1)
MAXIMUM MUTUAL INFORMATION TRAINING (1)
MEL FREQUENCY CEPSTRAL COEFFICIENT (1)
MINIMISATION (1)
MINIMUM DELETION ERROR (1)
MINIMUM INSERTION ERROR (1)
MINIMUM SUBSTITUTION ERROR (1)
MULTIPLE FRAME RATES (1)
MULTIPLE TIME RESOLUTION ANALYSIS (1)
MULTISCALE FEATURES (1)
MWE (1)
OBJECTIVE FUNCTION LEVEL (1)
OPTIMIZATION TECHNIQUE (1)
PARAMETER ESTIMATION (1)
PARTITION ACOUSTIC MODEL OPTIMIZATION (1)
PATTERN RECOGNITION (1)
PATTERN RECOGNITION APPLICATION (1)
PATTERN RECOGNTION (1)
PHONEME RECOGNITION (1)
PHONETIC RECOGNITION (1)
RATIONAL-FUNCTION FORM (1)
RATIONAL-FUNCTION OPTIMIZATION (1)
SEQUENTIAL PATTERN RECOGNITION (1)
SIGNAL CLASSIFICATION (1)
SPEECH RECOGNITION AND AUDIO PROCESSING (1)
SPEECH SIGNAL (1)
SUPPORT VECTOR MACHINES (1)
SVM (1)
TIMIT PHONE RECOGNITION TASK (1)
TRAINING CRITERIA (1)
VARIABLE FRAME RATE (1)
VOCABULARY (1)
VOCABULARY TASK (1)
WEIGHTED FINITE STATE TRANSDUCER (1)
WORD ERROR RATE (1)
more

INFONA - science communication portal

Search results

Modified MPE/MMI in a transducer-based framework

Multiple time resolution analysis of speech signal using MCE training with application to speech recognition

Discriminative learning in sequential pattern recognition

An enhanced minimum classification error learning framework for balancing insertion, deletion and substitution errors

Comparison of Large Margin Training to Other Discriminative Methods for Phonetic Recognition by Hidden Markov Models

A Novel Learning Method for Hidden Markov Models in Speech and Audio Processing

Add recipient

Sending message cancelled

Are you sure you want to cancel sending this message?

Send message

Filter options

Publication date

Date range setting

Set the date range to filter the displayed results. You can set a starting date, ending date or both. You can enter the dates manually or choose them from the calendar.

Publication type

Keywords

Reporting an error / abuse

Sending the report failed

Accessibility options