paper-with-me

홈 › Papers

Deep Cross Residual Learning for Multitask Visual Recognition

2016-04-05 · Brendan Jou, Shih-Fu Chang

Residual learning has recently surfaced as an effective means of constructing very deep neural networks for object recognition. However, current incarnations of residual networks do not allow for the modeling and integration of complex relations between closely coupled recognition tasks or across domains. Such problems are often encountered in multimedia applications involving large-scale content recognition. We propose a novel extension of residual learning for deep networks that enables intuitive learning across multiple related tasks using cross-connections called cross-residuals. These cross-residuals connections can be viewed as a form of in-network regularization and enables greater network generalization. We show how cross-residual learning (CRL) can be integrated in multitask networks to jointly train and detect visual concepts across several tasks. We present a single multitask cross-residual network with >40% less parameters that is able to achieve competitive, or even better, detection performance on a visual sentiment concept detection problem normally requiring multiple specialized single-task networks. The resulting multitask cross-residual network also achieves better detection performance by about 10.4% over a standard multitask residual network without cross-residuals with even a small amount of cross-task weighting.

📄 PDF Abstract BibTeX arXiv:1604.01335

Code (1)

imatge-upc/affective-2017-musa2 tf

Tasks

Object Recognition

Similar Papers 제목 키워드 기반

Multitask Learning and Multistage Fusion for Dimensional Audiovisual Emotion Recognition

2020-02-26 · ICASSP 2020 4 · Bagus Tris Atmaja, Masato Akagi

Due to its ability to accurately predict emotional state using multimodal features, audiovisual emotion recognition has recently gained more interest from researchers. This paper proposes two methods to predict emotional…

AttributeEmotion Recognition

Residual Learning Inspired Crossover Operator and Strategy Enhancements for Evolutionary Multitasking

2025-03-27 · Ruilin Wang, Xiang Feng, Huiqun Yu, Edmund M-K Lai

In evolutionary multitasking, strategies such as crossover operators and skill factor assignment are critical for effective knowledge transfer. Existing improvements to crossover operators primarily focus on low-dimensio…

Super-ResolutionTransfer Learning

Aligned Image-Word Representations Improve Inductive Transfer Across Vision-Language Tasks

2017-04-02 · ICCV 2017 10 · Tanmay Gupta, Kevin Shih, Saurabh Singh, Derek Hoiem

An important goal of computer vision is to build systems that learn visual representations over time that can be applied to many tasks. In this paper, we investigate a vision-language embedding as a core representation a…

Multi-Task LearningQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Multitask Identity-Aware Image Steganography via Minimax Optimization

2021-07-13 · Jiabao Cui, Pengyi Zhang, Songyuan Li, Liangli Zheng 외

High-capacity image steganography, aimed at concealing a secret image in a cover image, is a technique to preserve sensitive data, e.g., faces and fingerprints. Previous methods focus on the security during transmission …

Image RestorationImage Steganography

Learning from Annotation Uncertainty: Entropy-Aware Curriculum for Speech Emotion Recognition

2026-06-25 · Zahra Omidi, John H. L. Hansen arxiv

Speech emotion recognition (SER) often relies on hard consensus labels that collapse annotator disagreement. We study distribution-based supervision for 9-class SER on MSP-Podcast 2.0 using a WavLM-Base multitask model f…

Speech Emotion Recognition