Tze Yuang Chong

chapter

On the study of very low-resource language keyword search

Van Tung Pham, Haihua Xu, Van Hai Do, Tze Yuang Chong, more

2015 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA) > 358 - 364

2015 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA)

In this paper we report our approaches to accomplishing the very limited resource keyword search (KWS) task in the NIST Open Keyword Search 2015 (OpenKWS15) Evaluation. We devised the methods, first, to attain better acoustic modeling, multilingual and semi-supervised acoustic model training as well as the examplar-based acoustic model training; second, to address the overwhelming out-of-vocabulary...

article

Decoupling Word-Pair Distance and Co-occurrence Information for Effective Long History Context Language Modeling

Tze Yuang Chong, Rafael E. Banchs, Eng Siong Chng, Haizhou Li

IEEE/ACM Transactions on Audio, Speech, and Language Processing > 2015 > 23 > 7 > 1221 - 1232

In this paper, we propose the use of distance and co-occurrence information of word-pairs to improve language modeling. We have empirically shown that, for history-context sizes of up to ten words, the extracted information about distance and co-occurrence complements the $n$ -gram language model well, for which learning long-history contexts is inherently difficult. Evaluated on the Wall Street Journal...

chapter

Improving language modeling by using distance and co-occurrence information of word-pairs and its application to LVCSR

Tze Yuang Chong, Rafael E. Banchs, Eng Siong Chng, Haizhou Li

2014 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) > 4883 - 4887

ICASSP 2014 - 2014 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

This paper reports our study in exploiting the distance and co-occurrence information of word-pairs to improve the n-gram language model. We used these two types of information for modeling the distant context, up to history length of ten. Also we show that the proposed model provides complementary information about the n-gram's context that is unable to be captured by the n-gram model due to data...

chapter

The development and analysis of a Malay broadcasr news corpus

Tze Yuang Chong, Xiong Xiao, Haihua Xu, Tien-Ping Tan, more

2013 International Conference Oriental COCOSDA held jointly with 2013 Conference on Asian Spoken Language Research and Evaluation (O-COCOSDA/CASLRE) > 1 - 5

2013 International Conference Oriental COCOSDA held jointly with 2013 Conference on Asian Spoken Language Research and Evaluation (O-COCOSDA/CASLRE)

This paper presents our effort in collecting a Malay broadcast news (BN) speech corpus to support our research in Malay LVCSR. The 53 hours corpus is recorded from the TV channels in both Singapore and Malaysia over a 9-month period. To facilitate various researches in LVCSR, besides of orthographic transcription, the corpus provides other metadata such as speaking environment type, speaker identity...

chapter

Collection and annotation of Malay conversational speech corpus

Tze Yuang Chong, Xiong Xiao, Tien-Ping Tan, Eng Siong Chng, more

2012 International Conference on Speech Database and Assessments > 30 - 35

2012 Oriental COCOSDA 2012 - International Conference on Speech Database and Assessments

We report the development of a Malay conversational speech corpus as part of our research in spontaneous conversational speech LVCSR. This corpus development effort is the collaboration between NTU and USM. The goal is to collect, transcribe, and annotate 50 hours of conversational Malay speech. The conversation is recorded from both close-talk and telephone channels, and both speakers' utterances...

INFONA - science communication portal

Search results for: Tze Yuang Chong

On the study of very low-resource language keyword search

Decoupling Word-Pair Distance and Co-occurrence Information for Effective Long History Context Language Modeling

Improving language modeling by using distance and co-occurrence information of word-pairs and its application to LVCSR

The development and analysis of a Malay broadcasr news corpus

Collection and annotation of Malay conversational speech corpus

Filter options

Publication date

Publication type

Keywords

INFONA - science communication portal

Search results for: Tze Yuang Chong

On the study of very low-resource language keyword search

Decoupling Word-Pair Distance and Co-occurrence Information for Effective Long History Context Language Modeling

Improving language modeling by using distance and co-occurrence information of word-pairs and its application to LVCSR

The development and analysis of a Malay broadcasr news corpus

Collection and annotation of Malay conversational speech corpus

Add recipient

Sending message cancelled

Are you sure you want to cancel sending this message?

Send message

Filter options

Publication date

Date range setting

Set the date range to filter the displayed results. You can set a starting date, ending date or both. You can enter the dates manually or choose them from the calendar.

Publication type

Keywords

Reporting an error / abuse

Sending the report failed

Accessibility options