paper-with-me

Papers

Can machines learn to see without visual databases?

2021-10-12 · Alessandro Betti, Marco Gori, Stefano Melacci, Marcello Pelillo, Fabio Roli

This paper sustains the position that the time has come for thinking of learning machines that conquer visual skills in a truly human-like context, where a few human-like object supervisions are given by vocal interactions and pointing aids only. This likely requires new foundations on computational processes of vision with the final purpose of involving machines in tasks of visual description by living in their own visual environment under simple man-machine linguistic interactions. The challenge consists of developing machines that learn to see without needing to handle visual databases. This might open the doors to a truly orthogonal competitive track concerning deep learning technologies for vision which does not rely on the accumulation of huge visual databases.

📄 PDF Abstract BibTeX arXiv:2110.05973

Code (0)

등록된 구현이 없습니다.

Tasks

Position

Similar Papers 제목 키워드 기반

AdaGraph: Unifying Predictive and Continuous Domain Adaptation through Graphs

2019-03-17 · CVPR 2019 6 · Massimiliano Mancini, Samuel Rota Bulò, Barbara Caputo, Elisa Ricci

The ability to categorize is a cornerstone of visual intelligence, and a key functionality for artificial, autonomous visual machines. This problem will never be solved without algorithms able to adapt and generalize acr…

Domain Adaptation

A bagging and importance sampling approach to Support Vector Machines

2018-08-17 · R. Bárcenas, M. D. Gónzalez--Lima, A. J. Quiroz

An importance sampling and bagging approach to solving the support vector machine (SVM) problem in the context of large databases is presented and evaluated. Our algorithm builds on the nearest neighbors ideas presented …

Affect Analysis in-the-wild: Valence-Arousal, Expressions, Action Units and a Unified Framework

2021-03-29 · Dimitrios Kollias, Stefanos Zafeiriou

Affect recognition based on subjects' facial expressions has been a topic of major research in the attempt to generate machines that can understand the way subjects feel, act and react. In the past, due to the unavailabi…

Action Unit DetectionEmotion ClassificationFacial Action Unit Detection

Robust Deep Appearance Models

2016-07-03 · Kha Gia Quach, Chi Nhan Duong, Khoa Luu, Tien D. Bui

This paper presents a novel Robust Deep Appearance Models to learn the non-linear correlation between shape and texture of face images. In this approach, two crucial components of face images, i.e. shape and texture, are…

STAViS: Spatio-Temporal AudioVisual Saliency Network

2020-01-09 · CVPR 2020 6 · Antigoni Tsiami, Petros Koutras, Petros Maragos

We introduce STAViS, a spatio-temporal audiovisual saliency network that combines spatio-temporal visual and auditory information in order to efficiently address the problem of saliency estimation in videos. Our approach…

Saliency Prediction