paper-with-me

Papers

Tri-Select: A Multi-Stage Visual Data Selection Framework for Mobile Visual Crowdsensing

2025-12-18 · Jiayu Zhang, Kaixing Zhao, Tianhao Shao, Bin Guo, Liang He arxiv

Mobile visual crowdsensing enables large-scale, fine-grained environmental monitoring through the collection of images from distributed mobile devices. However, the resulting data is often redundant and heterogeneous due to overlapping acquisition perspectives, varying resolutions, and diverse user behaviors. To address these challenges, this paper proposes Tri-Select, a multi-stage visual data selection framework that efficiently filters redundant and low-quality images. Tri-Select operates in three stages: (1) metadata-based filtering to discard irrelevant samples; (2) spatial similarity-based spectral clustering to organize candidate images; and (3) a visual-feature-guided selection based on maximum independent set search to retain high-quality, representative images. Experiments on real-world and public datasets demonstrate that Tri-Select improves both selection efficiency and dataset quality, making it well-suited for scalable crowdsensing applications.

📄 PDF Abstract BibTeX arXiv:2512.16469

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning

2025-09-26 · Zejun Li, Yingxiu Zhao, Jiwen Zhang, Siyuan Wang 외 arxiv

Current visual reasoning methods mainly focus on exploring specific reasoning modes. Although improvements can be achieved in particular domains, they struggle to develop general reasoning capabilities. Inspired by this,…

Visual Reasoning

KTV: Keyframes and Key Tokens Selection for Efficient Training-Free Video LLMs

2026-02-03 · Baiyang Song, Jun Peng, Yuxin Zhang, Guangyao Chen 외 arxiv

Training-free video understanding leverages the strong image comprehension capabilities of pre-trained vision language models (VLMs) by treating a video as a sequence of static frames, thus obviating the need for costly …

Progressive Multi-Stage Learning for Discriminative Tracking

2020-04-01 · Weichao Li, Xi Li, Omar Elfarouk Bourahla, Fuxian Huang 외

Visual tracking is typically solved as a discriminative learning problem that usually requires high-quality samples for online model adaptation. It is a critical and challenging problem to evaluate the training samples c…

Visual Tracking

Interactive Visual Exploration of Latent Space (IVELS) for peptide auto-encoder model selection

2019-03-27 · ICLR Workshop DeepGenStruct 2019 · Tom Sercu, Sebastian Gehrmann, Hendrik Strobelt, Payel Das 외

We present a tool for Interactive Visual Exploration of Latent Space (IVELS) for model selection. Evaluating generative models of discrete sequences from a continuous latent space is a challenging problem, since…

Model Selection

Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers

2026-05-22 · Shuhong Zheng, Michael Oechsle, Erik Sandström, Marie-Julie Rakotosaona 외 arxiv

Visual geometry transformers have become powerful architectures for multi-view 3D reconstruction, enabling joint prediction of multiple 3D attributes in a feed-forward manner. However, their computational cost grows quad…

Multi-View 3D Reconstruction