paper-with-me

Papers

Investigating salient representations and label Variance in Dimensional Speech Emotion Analysis

2023-12-17 · Vikramjit Mitra, Jingping Nie, Erdrin Azemi

Representations derived from models such as BERT (Bidirectional Encoder Representations from Transformers) and HuBERT (Hidden units BERT), have helped to achieve state-of-the-art performance in dimensional speech emotion recognition. Despite their large dimensionality, and even though these representations are not tailored for emotion recognition tasks, they are frequently used to train large speech emotion models with high memory and computational costs. In this work, we show that there exist lower-dimensional subspaces within the these pre-trained representational spaces that offer a reduction in downstream model complexity without sacrificing performance on emotion estimation. In addition, we model label uncertainty in the form of grader opinion variance, and demonstrate that such information can improve the models generalization capacity and robustness. Finally, we compare the robustness of the emotion models against acoustic degradations and observed that the reduced dimensional representations were able to retain the performance similar to the full-dimensional representations without significant regression in dimensional emotion performance.

📄 PDF Abstract BibTeX arXiv:2312.16180

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion RecognitionSpeech Emotion Recognition

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Attention 설명 없음
Linear Warmup With Linear Decay Linear Warmup With Linear Decay is a learning rate schedule in which we increase the learning rate linearly for $n$ updates and then linearly decay afterwards.
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Multi-Head Attention 설명 없음
Adam 설명 없음

Similar Papers 제목 키워드 기반

H-SPLID: HSIC-based Saliency Preserving Latent Information Decomposition

2025-10-23 · Lukas Miklautz, Chengzhi Shi, Andrii Shkabrii, Theodoros Thirimachos Davarakis 외 arxiv

We introduce H-SPLID, a novel algorithm for learning salient feature representations through the explicit decomposition of salient and non-salient features into separate spaces. We show that H-SPLID promotes learning low…

Image Classification

Invariant properties of a locally salient dither pattern with a spatial-chromatic histogram

2018-02-28 · A. M. R. R. Bandara, L. Ranathunga, N. A. Abdullah

Compacted Dither Pattern Code (CDPC) is a recently found feature which is successful in irregular shapes based visual depiction. Locally salient dither pattern feature is an attempt to expand the capability of CDPC for b…

Salient Region Detection via High-Dimensional Color Transform

2014-06-01 · CVPR 2014 6 · Jiwhan Kim, Dongyoon Han, Yu-Wing Tai, Junmo Kim

In this paper, we introduce a novel technique to automatically detect salient regions of an image via high-dimensional color transform. Our main idea is to represent a saliency map of an image as a linear combination of …

Vocal Bursts Intensity Prediction

Salient Object Detection: A Discriminative Regional Feature Integration Approach

2014-10-22 · CVPR 2013 6 · Huaizu Jiang, Zejian yuan, Ming-Ming Cheng, Yihong Gong 외

Salient object detection has been attracting a lot of interest, and recently various heuristic computational models have been designed. In this paper, we formulate saliency map computation as a regression problem. Our me…

Image SegmentationObjectobject-detectionObject Detection+3

Why Do Self-Supervised Models Transfer? Investigating the Impact of Invariance on Downstream Tasks

2021-11-22 · Linus Ericsson, Henry Gouk, Timothy M. Hospedales

Self-supervised learning is a powerful paradigm for representation learning on unlabelled images. A wealth of effective new methods based on instance matching rely on data-augmentation to drive learning, and these have r…

Data AugmentationRepresentation LearningSelf-Supervised Learning