Search results

Items from 1 to 6 out of 6 results

chapter

Combining regression and classification methods for improving automatic speaker age recognition

C van Heerden, E Barnard, M Davel, C van der Walt, more

2010 IEEE International Conference on Acoustics, Speech and Signal Processing > 5174 - 5177

2010 IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP 2010

We present a novel approach to automatic speaker age classification, which combines regression and classification to achieve competitive classification accuracy on telephone speech. Support vector machine regression is used to generate finer age estimates, which are combined with the posterior probabilities of well-trained discriminative gender classifiers to predict both the age and gender of a speaker...

article

Comparison of Speaker Adaptation Methods as Feature Extraction for SVM-Based Speaker Recognition

M Ferras, Cheung-Chi Leung, C Barras, J.-L. Gauvain

IEEE Transactions on Audio, Speech, and Language Processing > 2010 > 18 > 6 > 1366 - 1378

In the last years the speaker recognition field has made extensive use of speaker adaptation techniques. Adaptation allows speaker model parameters to be estimated using less speech data than needed for maximum-likelihood (ML) training. The maximum a posteriori (MAP) and maximum-likelihood linear regression (MLLR) techniques have typically been used for adaptation. Recently, MAP and MLLR adaptation...

chapter

Exploiting prosodic information for Speaker Recognition

Yanhua Long, Bin Ma, Haizhou Li, Wu Guo, more

2009 IEEE International Conference on Acoustics, Speech and Signal Processing > 4225 - 4228

ICASSP 2009 - 2009 IEEE International Conference on Acoustics, Speech and Signal Processing

In this paper, we study speaker characterization using prosodic supervectors with negative within-class covariance normalization (NWCCN) projection and speaker modeling with support vector regression (SVR). We also propose a segmental weight fusion (SWF) technique that combines acoustic and prosodic subsystems effectively, despite the big performance gap between the subsystems. We validate the effectiveness...

chapter

SVM Based Speaker Recognition Using Maximum a posteriori Linear Regression

Xiang Zhang, Qingwei Zhao, Yonghong Yan

2009 International Conference on Electronic Computer Technology > 438 - 442

2009 International Conference on Electronic Computer Technology. ICECT 2009

Maximum likelihood linear regression (MLLR) is a widely used technique for speaker adaptation in large vocabulary speech recognition system. Recently, using MLLR transforms as features for SVM based speaker recognition tasks has been proposed, achieving performance comparable to that obtained with cepstral features. In this paper, we focus on calculating the transforms based on a GMM universal background...

chapter

Pitch Prediction from MFCC Vectors Using Support Vector Regression

Changping Peng, Wenjiu Liu

2007 International Conference on Natural Language Processing and Knowledge Engineering > 342 - 346

2007 IEEE International Conference on Natural Language Processing and Knowledge Engineering (NLP-KE '07)

Mel-frequency cepstral coefficients (MFCC) are proved to be the effective feature for speech recognition and speaker recognition, while pitch frequency is also one of the favorite prosodic features. This work manages to bridge them with hidden Markov model (HMM) and support vector regression (SVR). A set of speaker-independent HMMs is used to align the training data, so that the parameters of the...

chapter

Constrained MLLR for Speaker Recognition

M. Ferras, Cheung Chi Leung, C. Barras, J.-L. Gauvain

2007 IEEE International Conference on Acoustics, Speech and Signal Processing - ICASSP '7 > 4 > IV-53 - IV-56

2007 IEEE International Conference on Acoustics, Speech, and Signal Processing

One particularly difficult challenge for cross-channel MLLR (CMLLR) are two widely-used techniques for speaker introduced in the 2005 and 2006 NIST Speaker Recognition Evaluations, where training uses telephone speech and verification uses speech from multiple auxiliary comparable to that obtained with cepstral features. This paper describes a new feature extraction technique for speaker recognition...

Filter options

Keywords:
REGRESSION ANALYSIS
FEATURE EXTRACTION
SPEAKER RECOGNITION

Publication date

Set your own date range

INFONA - science communication portal

Search results

Combining regression and classification methods for improving automatic speaker age recognition

Comparison of Speaker Adaptation Methods as Feature Extraction for SVM-Based Speaker Recognition

Exploiting prosodic information for Speaker Recognition

SVM Based Speaker Recognition Using Maximum a posteriori Linear Regression

Pitch Prediction from MFCC Vectors Using Support Vector Regression

Constrained MLLR for Speaker Recognition

Add recipient

Sending message cancelled

Are you sure you want to cancel sending this message?

Send message

Filter options

Publication date

Date range setting

Set the date range to filter the displayed results. You can set a starting date, ending date or both. You can enter the dates manually or choose them from the calendar.

Content availability

Publication type

Keywords

Reporting an error / abuse

Sending the report failed

Accessibility options