paper-with-me

홈 › Papers

Efficient Pipelines for Vision-Based Context Sensing

2020-11-01 · Xiaochen Liu

Context awareness is an essential part of mobile and ubiquitous computing. Its goal is to unveil situational information about mobile users like locations and activities. The sensed context can enable many services like navigation, AR, and smarting shopping. Such context can be sensed in different ways including visual sensors. There is an emergence of vision sources deployed worldwide. The cameras could be installed on roadside, in-house, and on mobile platforms. This trend provides huge amount of vision data that could be used for context sensing. However, the vision data collection and analytics are still highly manual today. It is hard to deploy cameras at large scale for data collection. Organizing and labeling context from the data are also labor intensive. In recent years, advanced vision algorithms and deep neural networks are used to help analyze vision data. But this approach is limited by data quality, labeling effort, and dependency on hardware resources. In summary, there are three major challenges for today's vision-based context sensing systems: data collection and labeling at large scale, process large data volumes efficiently with limited hardware resources, and extract accurate context out of vision data. The thesis explores the design space that consists of three dimensions: sensing task, sensor types, and task locations. Our prior work explores several points in this design space. We make contributions by (1) developing efficient and scalable solutions for different points in the design space of vision-based sensing tasks; (2) achieving state-of-the-art accuracy in those applications; (3) and developing guidelines for designing such sensing systems.

📄 PDF Abstract BibTeX arXiv:2011.00427

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Leveraging Vision Reconstruction Pipelines for Satellite Imagery

2019-10-07 · Kai Zhang, Jin Sun, Noah Snavely

Reconstructing 3D geometry from satellite imagery is an important topic of research. However, disparities exist between how this 3D reconstruction problem is handled in the remote sensing context and how multi-view recon…

3D geometry3D Reconstruction

Improving Visual Grounding in Remote Sensing via Cluster-Guided Refinement and Model Ensemble Voting

2026-05-30 · Panav Shah, Geet Sethi, Ashutosh Gandhe arxiv

Visual grounding aims to locate image regions that correspond to natural language descriptions and is a key component of interpretable vision systems. In remote sensing imagery, grounding is particularly challenging due …

Visual Grounding

Revisiting Change VQA in Remote Sensing with Structured and Native Multimodal Qwen Models

2026-04-20 · Yakoub Bazi, Mohamad M. Al Rahhal, Mansour Zuair, Faroun Mohamed arxiv

Change visual question answering (Change VQA) addresses the problem of answering natural-language questions about semantic changes between bi-temporal remote sensing (RS) images. Although vision-language models (VLMs) ha…

Visual Question Answering

FlexiCup: Wireless Multimodal Suction Cup with Dual-Zone Vision-Tactile Sensing

2025-11-18 · Junhao Gong, Shoujie Li, Kit-Wa Sou, Changqing Guo 외 arxiv

Conventional suction cups lack sensing capabilities for contact-aware manipulation in unstructured environments. This paper presents FlexiCup, a multimodal suction cup with wireless electronics that integrate dual-zone v…

SatBLIP: Context Understanding and Feature Identification from Satellite Imagery with Vision-Language Learning

2026-04-15 · Xue Wu, Shengting Cao, Shenglin Li, Jiaqi Gong arxiv

Rural environmental risks are shaped by place-based conditions (e.g., housing quality, road access, land-surface patterns), yet standard vulnerability indices are coarse and provide limited insight into risk contexts. We…