paper-with-me

홈 › Papers

VISIR: Visual and Semantic Image Label Refinement

2019-09-02 · Sreyasi Nag Chowdhury, Niket Tandon, Hakan Ferhatosmanoglu, Gerhard Weikum

The social media explosion has populated the Internet with a wealth of images. There are two existing paradigms for image retrieval: 1) content-based image retrieval (CBIR), which has traditionally used visual features for similarity search (e.g., SIFT features), and 2) tag-based image retrieval (TBIR), which has relied on user tagging (e.g., Flickr tags). CBIR now gains semantic expressiveness by advances in deep-learning-based detection of visual labels. TBIR benefits from query-and-click logs to automatically infer more informative labels. However, learning-based tagging still yields noisy labels and is restricted to concrete objects, missing out on generalizations and abstractions. Click-based tagging is limited to terms that appear in the textual context of an image or in queries that lead to a click. This paper addresses the above limitations by semantically refining and expanding the labels suggested by learning-based object detection. We consider the semantic coherence between the labels for different objects, leverage lexical and commonsense knowledge, and cast the label assignment into a constrained optimization problem solved by an integer linear program. Experiments show that our method, called VISIR, improves the quality of the state-of-the-art visual labeling tools like LSDA and YOLO.

📄 PDF Abstract BibTeX arXiv:1909.00741

Code (0)

등록된 구현이 없습니다.

Tasks

Content-Based Image RetrievalImage Retrievalobject-detectionObject DetectionRetrievalTAG

Similar Papers 제목 키워드 기반

ViSIR: Vision Transformer Single Image Reconstruction Method for Earth System Models

2025-02-10 · Ehsan Zeraatkar, Salah Faroughi, Jelena Tešić

Purpose: Earth system models (ESMs) integrate the interactions of the atmosphere, ocean, land, ice, and biosphere to estimate the state of regional and global climate under a wide variety of conditions. The ESMs are high…

Image ReconstructionSSIM

JarvisIR: Elevating Autonomous Driving Perception with Intelligent Image Restoration

2025-01-01 · CVPR 2025 1 · Yunlong Lin, Zixu Lin, Haoyu Chen, Panwang Pan 외

Vision-centric perception systems often struggle with unpredictable and coupled weather degradations in the wild. Current solutions are often limited, as they either depend on specific degradation priors or suffer fr…

Autonomous DrivingImage Restoration

Incremental Image Labeling via Iterative Refinement

2023-04-18 · Fausto Giunchiglia, Xiaolei Diao, Mayukh Bagchi

Data quality is critical for multimedia tasks, while various types of systematic flaws are found in image benchmark datasets, as discussed in recent work. In particular, the existence of the semantic gap problem leads to…

Weakly Supervised Semantic Segmentation for Social Images

2015-06-01 · CVPR 2015 6 · Wei Zhang, Sheng Zeng, Dequan Wang, xiangyang xue

Image semantic segmentation is the task of partitioning image into several regions based on semantic concepts. In this paper, we learn a weakly supervised semantic segmentation model from social images whose labels a…

SegmentationSemantic SegmentationWeakly supervised Semantic SegmentationWeakly-Supervised Semantic Segmentation

Active Label Refinement for Semantic Segmentation of Satellite Images

2023-09-12 · Tuan Pham Minh, Jayan Wijesingha, Daniel Kottke, Marek Herde 외

Remote sensing through semantic segmentation of satellite images contributes to the understanding and utilisation of the earth's surface. For this purpose, semantic segmentation networks are typically trained on large se…

Active LearningSegmentationSemantic Segmentation