Wyniki wyszukiwania dla: S. Araki

Pozycje od 1 do 5 spośród 5 wyników

rozdział

Online meeting recognizer with multichannel speaker diarization

S Araki, T Hori, M Fujimoto, S Watanabe, więcej

2010 Conference Record of the Forty Fourth Asilomar Conference on Signals, Systems and Computers > 1697 - 1701

2010 44th Asilomar Conference on Signals, Systems and Computers

We present our newly developed real-time conversation analyzer for group meetings. The goal of the system is to estimate automatically “who speaks when and what” in an online manner. In our system, “who speaks when” information is first obtained by estimating the directions of arrival (DOAs) of signals. Then, “who speaks what” is estimated with our automatic speech recognition (ASR) system, after...

artykuł

Speech Activity Detection for Multi-Party Conversation Analyses Based on Likelihood Ratio Test on Spatial Magnitude

K Ishizuka, S Araki, T Kawahara

IEEE Transactions on Audio, Speech, and Language Processing > 2010 > 18 > 6 > 1354 - 1365

This paper proposes a microphone array-based speech activity detection (SAD) method for analyzing multi-party conversations recorded in the presence of noise. In particular, the proposed method considers conversations where the number of speakers and speaker locations cannot be restricted, such as when standing and talking, and at poster sessions. When we observe such conversations, there are directional...

rozdział

A probabilistic speaker clustering for DOA-based diarization

K. Ishiguro, T. Yamada, S. Araki, T. Nakatani

2009 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics > 241 - 244

2009 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA)

We present a probabilistic speaker clustering and diarization model. Speaker diarization determines ldquowho spoke whenrdquo from the recorded conversation of unknown number of people. We formulate this problem as the clustering of sequential auditory features generated by an unknown number of latent mixture components (speakers). We employ a probabilistic model which automatically estimates the number...

rozdział

A DOA Based Speaker Diarization System for Real Meetings

S. Araki, M. Fujimoto, K. Ishizuka, H. Sawada, więcej

2008 Hands-Free Speech Communication and Microphone Arrays > 29 - 32

2008 Hands-Free Speech Communication and Microphone Arrays (HSCMA '08)

This paper presents a speaker diarization system that estimates who spoke when in a meeting. Our proposed system is realized by using a noise robust voice activity detector (VAD), a direction of arrival (DOA) estimator, and a DOA classifier. Our previous system utilized the generalized cross correlation method with the phase transform (GCC-PHAT) approach for the DOA estimation. Because the GCC-PHAT...

rozdział

Speaker indexing and speech enhancement in real meetings / conversations

S. Araki, M. Fujimoto, K. Ishizuka, H. Sawada, więcej

2008 IEEE International Conference on Acoustics, Speech and Signal Processing > 93 - 96

ICASSP 2008. IEEE International Conference on Acoustic, Speech and Signal Processes

This paper presents a speaker indexing method that uses a small number of microphones to estimate who spoke when. Our proposed speaker indexing is realized by using a noise robust voice activity detector (VAD), a QCC-PHAT based direction of arrival (DOA) estimator, and a DOA classifier. Using the estimated speaker indexing information, we can also enhance the utterances of each speaker with a maximum...

Opcje filtrowania

Słowa kluczowe:
SPEAKER RECOGNITION

Data publikacji

Ustaw własny zakres dat

Typ publikacji

książka (4)
artykuł (1)

Słowa kluczowe

DIRECTION OF ARRIVAL ESTIMATION (3)
SPEECH ENHANCEMENT (3)
CONVERSATION RECORDING (2)
DIARIZATION (2)
DIRECTION OF ARRIVAL (2)
MICROPHONES (2)
SPEECH (2)
TIME 350 MS (2)
TIME-FREQUENCY ANALYSIS (2)
VOICE ACTIVITY DETECTOR (2)
ADAPTATION MODEL (1)
ASR SYSTEM (1)
AUDIO RECORDING (1)
AUDIO RECORDINGS (1)
AUTOMATIC SPEECH RECOGNITION SYSTEM (1)
BACKGROUND NOISE (1)
COMPUTATIONAL MODELING (1)
CROSSPOWER SPECTRUM PHASE (1)
DIARIZATION ERROR RATE (1)
DIFFUSE NOISE (1)
DIRECTION OF ARRIVAL ESTIMATIONS (1)
DIRECTION OF ARRIVAL ESTIMATOR (1)
DIRECTIONAL NOISE SOURCES (1)
DIRECTIONS OF ARRIVAL ESTIMATION (1)
DOA CLASSIFIER (1)
DOA ESTIMATION (1)
DOA-BASED DIARIZATION (1)
FEATURE EXTRACTION (1)
GENERALIZED CROSS CORRELATION METHOD (1)
HIDDEN MARKOV MODELS (1)
INDEXES (1)
INDEXING (1)
INSERTION ERROR REDUCTION (1)
INTERFERENCE SPEAKER VOICE SUPPRESSION (1)
LIGHT RAIL SYSTEMS (1)
LIKELIHOOD RATIO TEST BASED SAD METHOD (1)
MAGNITUDE COHERENCE (1)
MAXIMUM SNR BEAMFORMER (1)
MICROPHONE ARRAY-BASED SPEECH ACTIVITY DETECTION METHOD (1)
MICROPHONE ARRAYS (1)
MULTI-PARTY CONVERSATIONS (1)
MULTICHANNEL SPEAKER DIARIZATION (1)
MULTIPARTY CONVERSATION ANALYSIS (1)
NOISE (1)
NOISE ROBUST VOICE ACTIVITY DETECTOR (1)
OBSERVED SPECTRA (1)
ONLINE MEETING RECOGNIZER (1)
PHASE TRANSFORM (1)
PROBABILISTIC CLUSTERING (1)
PROBABILISTIC LOGIC (1)
PROBABILISTIC MODEL (1)
PROBABILISTIC SPEAKER CLUSTERING (1)
PROBABILITY (1)
REAL MEETING (1)
REAL RECORDED MEETINGS (1)
REAL-TIME CONVERSATION ANALYZER (1)
REVERBERATION (1)
REVERBERATION SUPPRESSION (1)
REVERBERATION TIME (1)
SIGNAL CLASSIFICATION (1)
SIGNAL DENOISING (1)
SIGNAL DETECTION (1)
SIGNAL TO NOISE RATIO (1)
SPATIAL INFORMATION (1)
SPATIAL MAGNITUDE (1)
SPEAKER DIARIZATION (1)
SPEAKER DIARIZATION SYSTEM (1)
SPEAKER INDEXING (1)
SPEECH ACTIVITY DETECTION (1)
SPEECH ACTIVITY DETECTION (SAD) (1)
SPEECH ANALYSIS (1)
SPEECH CODING (1)
SPEECH RECOGNITION (1)
TARGET SPEECH SIGNALS (1)
TESTING (1)
TIME FREQUENCY ANALYSIS (1)
TIME-FREQUENCY MASKING ESTIMATION (1)
TIME-FREQUENCY SLOT DOA ESTIMATION (1)
TIME-VARYING SPEAKER PROPORTION (1)
TIME-VARYING SYSTEMS (1)
VARIATIONAL BAYES (1)
WORKING ENVIRONMENT NOISE (1)
więcej

INFONA - portal komunikacji naukowej

Wyniki wyszukiwania dla: S. Araki

Online meeting recognizer with multichannel speaker diarization

Speech Activity Detection for Multi-Party Conversation Analyses Based on Likelihood Ratio Test on Spatial Magnitude

A probabilistic speaker clustering for DOA-based diarization

A DOA Based Speaker Diarization System for Real Meetings

Speaker indexing and speech enhancement in real meetings / conversations

Dodaj adresata

Anulowanie wysłania wiadomości

Czy na pewno chcesz anulować wysłanie wiadomości?

Wyślij wiadomość

Opcje filtrowania

Data publikacji

Ustawianie zakresu dat

Podaj zakres dat dla filtrowania wyświetlonych wyników. Możesz podać datę początkową, końcową lub obie daty. Daty możesz wpisać ręcznie lub wybrać za pomocą kalendarza.

Typ publikacji

Słowa kluczowe

Zgłaszanie błędu / nadużycia

Nieudane wysłanie zgłoszenia

Ułatwienia dostępu