Abstract
When Hidden Markov Models (HMMs) were first introduced, two competing representation models were proposed, the Moore model, with separate emission and transition distributions, which is commonly used in speech technologies, and the Mealy model, with a single emission-transition distribution. Since then the literature has mostly focused on the Moore model. In this paper, we would like to show the use of Mealy- HMMs for telephone conversation speaker diarization task. We present the Viterbi training and decoding for Mealy-HMMs and show that it yields similar performance compared to Moore- HMMs with a fewer number of parameters.
| Original language | English |
|---|---|
| Pages | 173-178 |
| Number of pages | 6 |
| State | Published - 1 Jan 2014 |
| Externally published | Yes |
| Event | Speaker and Language Recognition Workshop, Odyssey 2014 - Joensuu, Finland Duration: 16 Jun 2014 → 19 Jun 2014 |
Conference
| Conference | Speaker and Language Recognition Workshop, Odyssey 2014 |
|---|---|
| Country/Territory | Finland |
| City | Joensuu |
| Period | 16/06/14 → 19/06/14 |
ASJC Scopus subject areas
- Signal Processing
- Software
- Human-Computer Interaction
Fingerprint
Dive into the research topics of 'Telephone conversation speaker diarization using mealy-HMMs'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver