Search results for: S. Araki

Items from 1 to 5 out of 5 results

chapter

Online meeting recognizer with multichannel speaker diarization

S Araki, T Hori, M Fujimoto, S Watanabe, more

2010 Conference Record of the Forty Fourth Asilomar Conference on Signals, Systems and Computers > 1697 - 1701

2010 44th Asilomar Conference on Signals, Systems and Computers

We present our newly developed real-time conversation analyzer for group meetings. The goal of the system is to estimate automatically “who speaks when and what” in an online manner. In our system, “who speaks when” information is first obtained by estimating the directions of arrival (DOAs) of signals. Then, “who speaks what” is estimated with our automatic speech recognition (ASR) system, after...

article

Speech Activity Detection for Multi-Party Conversation Analyses Based on Likelihood Ratio Test on Spatial Magnitude

K Ishizuka, S Araki, T Kawahara

IEEE Transactions on Audio, Speech, and Language Processing > 2010 > 18 > 6 > 1354 - 1365

This paper proposes a microphone array-based speech activity detection (SAD) method for analyzing multi-party conversations recorded in the presence of noise. In particular, the proposed method considers conversations where the number of speakers and speaker locations cannot be restricted, such as when standing and talking, and at poster sessions. When we observe such conversations, there are directional...

chapter

A probabilistic speaker clustering for DOA-based diarization

K. Ishiguro, T. Yamada, S. Araki, T. Nakatani

2009 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics > 241 - 244

2009 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA)

We present a probabilistic speaker clustering and diarization model. Speaker diarization determines ldquowho spoke whenrdquo from the recorded conversation of unknown number of people. We formulate this problem as the clustering of sequential auditory features generated by an unknown number of latent mixture components (speakers). We employ a probabilistic model which automatically estimates the number...

chapter

A DOA Based Speaker Diarization System for Real Meetings

S. Araki, M. Fujimoto, K. Ishizuka, H. Sawada, more

2008 Hands-Free Speech Communication and Microphone Arrays > 29 - 32

2008 Hands-Free Speech Communication and Microphone Arrays (HSCMA '08)

This paper presents a speaker diarization system that estimates who spoke when in a meeting. Our proposed system is realized by using a noise robust voice activity detector (VAD), a direction of arrival (DOA) estimator, and a DOA classifier. Our previous system utilized the generalized cross correlation method with the phase transform (GCC-PHAT) approach for the DOA estimation. Because the GCC-PHAT...

chapter

Speaker indexing and speech enhancement in real meetings / conversations

S. Araki, M. Fujimoto, K. Ishizuka, H. Sawada, more

2008 IEEE International Conference on Acoustics, Speech and Signal Processing > 93 - 96

ICASSP 2008. IEEE International Conference on Acoustic, Speech and Signal Processes

This paper presents a speaker indexing method that uses a small number of microphones to estimate who spoke when. Our proposed speaker indexing is realized by using a noise robust voice activity detector (VAD), a QCC-PHAT based direction of arrival (DOA) estimator, and a DOA classifier. Using the estimated speaker indexing information, we can also enhance the utterances of each speaker with a maximum...

Filter options

Keywords:
SPEAKER RECOGNITION

Publication date

Set your own date range

Publication type

book (4)
article (1)

Keywords

DIRECTION OF ARRIVAL ESTIMATION (3)
SPEECH ENHANCEMENT (3)
CONVERSATION RECORDING (2)
DIARIZATION (2)
DIRECTION OF ARRIVAL (2)
MICROPHONES (2)
SPEECH (2)
TIME 350 MS (2)
TIME-FREQUENCY ANALYSIS (2)
VOICE ACTIVITY DETECTOR (2)
ADAPTATION MODEL (1)
ASR SYSTEM (1)
AUDIO RECORDING (1)
AUDIO RECORDINGS (1)
AUTOMATIC SPEECH RECOGNITION SYSTEM (1)
BACKGROUND NOISE (1)
COMPUTATIONAL MODELING (1)
CROSSPOWER SPECTRUM PHASE (1)
DIARIZATION ERROR RATE (1)
DIFFUSE NOISE (1)
DIRECTION OF ARRIVAL ESTIMATIONS (1)
DIRECTION OF ARRIVAL ESTIMATOR (1)
DIRECTIONAL NOISE SOURCES (1)
DIRECTIONS OF ARRIVAL ESTIMATION (1)
DOA CLASSIFIER (1)
DOA ESTIMATION (1)
DOA-BASED DIARIZATION (1)
FEATURE EXTRACTION (1)
GENERALIZED CROSS CORRELATION METHOD (1)
HIDDEN MARKOV MODELS (1)
INDEXES (1)
INDEXING (1)
INSERTION ERROR REDUCTION (1)
INTERFERENCE SPEAKER VOICE SUPPRESSION (1)
LIGHT RAIL SYSTEMS (1)
LIKELIHOOD RATIO TEST BASED SAD METHOD (1)
MAGNITUDE COHERENCE (1)
MAXIMUM SNR BEAMFORMER (1)
MICROPHONE ARRAY-BASED SPEECH ACTIVITY DETECTION METHOD (1)
MICROPHONE ARRAYS (1)
MULTI-PARTY CONVERSATIONS (1)
MULTICHANNEL SPEAKER DIARIZATION (1)
MULTIPARTY CONVERSATION ANALYSIS (1)
NOISE (1)
NOISE ROBUST VOICE ACTIVITY DETECTOR (1)
OBSERVED SPECTRA (1)
ONLINE MEETING RECOGNIZER (1)
PHASE TRANSFORM (1)
PROBABILISTIC CLUSTERING (1)
PROBABILISTIC LOGIC (1)
PROBABILISTIC MODEL (1)
PROBABILISTIC SPEAKER CLUSTERING (1)
PROBABILITY (1)
REAL MEETING (1)
REAL RECORDED MEETINGS (1)
REAL-TIME CONVERSATION ANALYZER (1)
REVERBERATION (1)
REVERBERATION SUPPRESSION (1)
REVERBERATION TIME (1)
SIGNAL CLASSIFICATION (1)
SIGNAL DENOISING (1)
SIGNAL DETECTION (1)
SIGNAL TO NOISE RATIO (1)
SPATIAL INFORMATION (1)
SPATIAL MAGNITUDE (1)
SPEAKER DIARIZATION (1)
SPEAKER DIARIZATION SYSTEM (1)
SPEAKER INDEXING (1)
SPEECH ACTIVITY DETECTION (1)
SPEECH ACTIVITY DETECTION (SAD) (1)
SPEECH ANALYSIS (1)
SPEECH CODING (1)
SPEECH RECOGNITION (1)
TARGET SPEECH SIGNALS (1)
TESTING (1)
TIME FREQUENCY ANALYSIS (1)
TIME-FREQUENCY MASKING ESTIMATION (1)
TIME-FREQUENCY SLOT DOA ESTIMATION (1)
TIME-VARYING SPEAKER PROPORTION (1)
TIME-VARYING SYSTEMS (1)
VARIATIONAL BAYES (1)
WORKING ENVIRONMENT NOISE (1)
more

INFONA - science communication portal

Search results for: S. Araki

Online meeting recognizer with multichannel speaker diarization

Speech Activity Detection for Multi-Party Conversation Analyses Based on Likelihood Ratio Test on Spatial Magnitude

A probabilistic speaker clustering for DOA-based diarization

A DOA Based Speaker Diarization System for Real Meetings

Speaker indexing and speech enhancement in real meetings / conversations

Add recipient

Sending message cancelled

Are you sure you want to cancel sending this message?

Send message

Filter options

Publication date

Date range setting

Set the date range to filter the displayed results. You can set a starting date, ending date or both. You can enter the dates manually or choose them from the calendar.

Publication type

Keywords

Reporting an error / abuse

Sending the report failed

Accessibility options