paper-with-me

Papers

Unsupervised Latent Behavior Manifold Learning from Acoustic Features: audio2behavior

2017-01-12 · Haoqi Li, Brian Baucom, Panayiotis Georgiou

Behavioral annotation using signal processing and machine learning is highly dependent on training data and manual annotations of behavioral labels. Previous studies have shown that speech information encodes significant behavioral information and be used in a variety of automated behavior recognition tasks. However, extracting behavior information from speech is still a difficult task due to the sparseness of training data coupled with the complex, high-dimensionality of speech, and the complex and multiple information streams it encodes. In this work we exploit the slow varying properties of human behavior. We hypothesize that nearby segments of speech share the same behavioral context and hence share a similar underlying representation in a latent space. Specifically, we propose a Deep Neural Network (DNN) model to connect behavioral context and derive the behavioral manifold in an unsupervised manner. We evaluate the proposed manifold in the couples therapy domain and also provide examples from publicly available data (e.g. stand-up comedy). We further investigate training within the couples' therapy domain and from movie data. The results are extremely encouraging and promise improved behavioral quantification in an unsupervised manner and warrants further investigation in a range of applications.

📄 PDF Abstract BibTeX arXiv:1701.03198

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DeepDiffusion: Unsupervised Learning of Retrieval-adapted Representations via Diffusion-based Ranking on Latent Feature Manifold

2021-12-14 · Takahiko Furuya, Ryutarou Ohbuchi

Unsupervised learning of feature representations is a challenging yet important problem for analyzing a large collection of multimedia data that do not have semantic labels. Recently proposed neural network-based unsuper…

Retrieval

From Parameters to Behaviors: Unsupervised Compression of the Policy Space

2025-09-26 · Davide Tenedini, Riccardo Zamboni, Mirco Mutti, Marcello Restelli arxiv

Despite its recent successes, Deep Reinforcement Learning (DRL) is notoriously sample-inefficient. We argue that this inefficiency stems from the standard practice of optimizing policies directly in the high-dimensional …

Reinforcement LearningContinuous Control

WavCube: Unifying Speech Representation for Understanding and Generation via Semantic-Acoustic Joint Modeling

2026-05-07 · Guanrou Yang, Tian Tan, Qian Chen, Zhikang Niu 외 arxiv

Integrating speech understanding and generation is a pivotal step toward building unified speech models. However, the different representations required for these two tasks currently pose significant compatibility challe…

Self-Supervised LearningSpeech EnhancementVoice Conversion

Self-Expressing Autoencoders for Unsupervised Spoken Term Discovery

2020-07-26 · Saurabhchand Bhati, Jesús Villalba, Piotr Żelasko, Najim Dehak

Unsupervised spoken term discovery consists of two tasks: finding the acoustic segment boundaries and labeling acoustically similar segments with the same labels. We perform segmentation based on the assumption that the …

Segmentation

Manifold GPLVMs for discovering non-Euclidean latent structure in neural data

2020-06-12 · NeurIPS 2020 12 · Kristopher T. Jensen, Ta-Chu Kao, Marco Tripodi, Guillaume Hennequin

A common problem in neuroscience is to elucidate the collective neural representations of behaviorally important variables such as head direction, spatial location, upcoming movements, or mental spatial transformations. …

Variational Inference