paper-with-me

Papers

Deep Saliency Prior for Reducing Visual Distraction

2021-09-05 · CVPR 2022 1 · Kfir Aberman, Junfeng He, Yossi Gandelsman, Inbar Mosseri, David E. Jacobs, Kai Kohlhoff, Yael Pritch, Michael Rubinstein

Using only a model that was trained to predict where people look at images, and no additional training data, we can produce a range of powerful editing effects for reducing distraction in images. Given an image and a mask specifying the region to edit, we backpropagate through a state-of-the-art saliency model to parameterize a differentiable editing operator, such that the saliency within the masked region is reduced. We demonstrate several operators, including: a recoloring operator, which learns to apply a color transform that camouflages and blends distractors into their surroundings; a warping operator, which warps less salient image regions to cover distractors, gradually collapsing objects into themselves and effectively removing them (an effect akin to inpainting); a GAN operator, which uses a semantic prior to fully replace image regions with plausible, less salient alternatives. The resulting effects are consistent with cognitive research on the human visual system (e.g., since color mismatch is salient, the recoloring operator learns to harmonize objects' colors with their surrounding to reduce their saliency), and, importantly, are all achieved solely through the guidance of the pretrained saliency model, with no additional supervision. We present results on a variety of natural images and conduct a perceptual study to evaluate and validate the changes in viewers' eye-gaze between the original images and our edited results.

📄 PDF Abstract BibTeX arXiv:2109.01980

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Predicting Visual Attention and Distraction During Visual Search Using Convolutional Neural Networks

2022-10-27 · Manoosh Samiei, James J. Clark

Most studies in computational modeling of visual attention encompass task-free observation of images. Free-viewing saliency considers limited scenarios of daily life. Most visual activities are goal-oriented and demand a…

A Saliency-Guided Street View Image Inpainting Framework for Efficient Last-Meters Wayfinding

2022-05-14 · Chuanbo Hu, Shan Jia, Fan Zhang, Xin Li

Global Positioning Systems (GPS) have played a crucial role in various navigation applications. Nevertheless, localizing the perfect destination within the last few meters remains an important but unresolved problem. Lim…

Image Inpaintingobject-detectionObject DetectionSalient Object Detection

Distraction is All You Need for Multimodal Large Language Model Jailbreaking

2025-02-15 · CVPR 2025 1 · Zuopeng Yang, Jiluan Fan, Anli Yan, Erdun Gao 외

Multimodal Large Language Models (MLLMs) bridge the gap between visual and textual data, enabling a range of advanced applications. However, complex internal interactions among visual elements and their alignment with te…

AllLanguage ModelingLanguage ModellingLarge Language Model+1

CAMERAS: Enhanced Resolution And Sanity preserving Class Activation Mapping for image saliency

2021-06-20 · CVPR 2021 1 · Mohammad A. A. K. Jalwana, Naveed Akhtar, Mohammed Bennamoun, Ajmal Mian

Backpropagation image saliency aims at explaining model predictions by estimating model-centric importance of individual pixels in the input. However, class-insensitivity of the earlier layers in a network only allows sa…

Learning Task Informed Abstractions

2021-03-09 · ICLR Workshop SSL-RL 2021 5 · Anonymous

Current model-based reinforcement learning methods struggle when operating from complex visual scenes due to their inability to prioritize task-relevant features. To mitigate this problem, we propose learning Task Inform…

Model-based Reinforcement Learningreinforcement-learningReinforcement Learning (RL)