paper-with-me

Papers

Normalization Before Shaking Toward Learning Symmetrically Distributed Representation Without Margin in Speech Emotion Recognition

2018-08-02 · Che-Wei Huang, Shrikanth. S. Narayanan

Regularization is crucial to the success of many practical deep learning models, in particular in a more often than not scenario where there are only a few to a moderate number of accessible training samples. In addition to weight decay, data augmentation and dropout, regularization based on multi-branch architectures, such as Shake-Shake regularization, has been proven successful in many applications and attracted more and more attention. However, beyond model-based representation augmentation, it is unclear how Shake-Shake regularization helps to provide further improvement on classification tasks, let alone the baffling interaction between batch normalization and shaking. In this work, we present our investigation on Shake-Shake regularization, drawing connections to the vicinal risk minimization principle and discriminative feature learning in verification tasks. Furthermore, we identify a strong resemblance between batch normalized residual blocks and batch normalized recurrent neural networks, where both of them share a similar convergence behavior, which could be mitigated by a proper initialization of batch normalization. Based on the findings, our experiments on speech emotion recognition demonstrate simultaneously an improvement on the classification accuracy and a reduction on the generalization gap both with statistical significance.

📄 PDF Abstract BibTeX arXiv:1808.00876

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationEmotion RecognitionGeneral ClassificationSpeech Emotion Recognition

Methods 이 논문이 사용한 방법론

Batch Normalization 설명 없음

Similar Papers 제목 키워드 기반

Linguistic Convergence in Societies with Asymmetrically Distributed Reputation

2014-07-01 · JEPTALNRECITAL 2014 7 · Gemma Bel-Enguix

Training Deep Neural Networks Without Batch Normalization

2020-08-18 · Divya Gaur, Joachim Folz, Andreas Dengel

Training neural networks is an optimization problem, and finding a decent set of parameters through gradient descent can be a difficult task. A host of techniques has been developed to aid this process before and during …

Tanh Works Better with Asymmetry

2023-09-21 · NeurIPS 2023 11

Batch Normalization is commonly located in front of activation functions, as proposed by the original paper. Swapping the order, i.e., using Batch Normalization after activation functions, has also been attempted, but it…

Overcoming Obstructions via Bandwidth-Limited Multi-Agent Spatial Handshaking

2021-07-01 · Nathaniel Glaser, Yen-Cheng Liu, Junjiao Tian, Zsolt Kira

In this paper, we address bandwidth-limited and obstruction-prone collaborative perception, specifically in the context of multi-agent semantic segmentation. This setting presents several key challenges, including proces…

Semantic Segmentation

Unsupervised Text Normalization Using Distributed Representations of Words and Phrases

2015-06-01 · WS 2015 6 · Vivek Kumar Rangarajan Sridhar
Entity Extraction using GANMachine TranslationSpeech RecognitionSpelling Correction+1