Exploration of A Self-Supervised Speech Model: A Study on Emotional Corpora

Li, Yuanchao
Mohamied, Yumnah
Bell, Peter
Lai, Catherine

Publication date

December 2022

Language

English

Abstract

Self-supervised speech models have grown fast during the past few years and have proven feasible for use in various downstream tasks. Some recent work has started to look at the characteristics of these models, yet many concerns have not been fully addressed. In this work, we conduct a study on emotional corpora to explore a popular self-supervised model -- wav2vec 2.0. Via a set of quantitative analysis, we mainly demonstrate that: 1) wav2vec 2.0 appears to discard paralinguistic information that is less useful for word recognition purposes; 2) for emotion recognition, representations from the middle layer alone perform as well as those derived from layer averaging, while the final layer results in the worst performance in some cases; 3) c...

Extracted data

We use cookies to provide a better user experience.

Data Protection

Exploration of A Self-Supervised Speech Model: A Study on Emotional Corpora

Abstract

Extracted data

Exploration of A Self-Supervised Speech Model: A Study on Emotional Corpora

Abstract

Extracted data

Related items

Related items