pith. sign in

arxiv: 1511.06772 · v1 · pith:KLFJX2WAnew · submitted 2015-11-20 · 📊 stat.ML

PLDA with Two Sources of Inter-session Variability

classification 📊 stat.ML
keywords variabilitymodelpldacapturechannelsconversationinter-sessionmultiple
0
0 comments X
read the original abstract

In some speaker recognition scenarios we find conversations recorded simultaneously over multiple channels. That is the case of the interviews in the NIST SRE dataset. To take advantage of that, we propose a modification of the PLDA model that considers two different inter-session variability terms. The first term is tied between all the recordings belonging to the same conversation whereas the second is not. Thus, the former mainly intends to capture the variability due to the phonetic content of the conversation while the latter tries to capture the channel variability. In this document, we derive the equations for this model. This model was applied in the paper "Handling Recordings Acquired Simultaneously over Multiple Channels with PLDA" published at Interspeech 2013.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.