paper-with-me

Papers

Minimalist Vision with Freeform Pixels

2024-12-30 · Jeremy Klotz, Shree K. Nayar

A minimalist vision system uses the smallest number of pixels needed to solve a vision task. While traditional cameras use a large grid of square pixels, a minimalist camera uses freeform pixels that can take on arbitrary shapes to increase their information content. We show that the hardware of a minimalist camera can be modeled as the first layer of a neural network, where the subsequent layers are used for inference. Training the network for any given task yields the shapes of the camera's freeform pixels, each of which is implemented using a photodetector and an optical mask. We have designed minimalist cameras for monitoring indoor spaces (with 8 pixels), measuring room lighting (with 8 pixels), and estimating traffic flow (with 8 pixels). The performance demonstrated by these systems is on par with a traditional camera with orders of magnitude more pixels. Minimalist vision has two major advantages. First, it naturally tends to preserve the privacy of individuals in the scene since the captured information is inadequate for extracting visual details. Second, since the number of measurements made by a minimalist camera is very small, we show that it can be fully self-powered, i.e., function without an external power supply or a battery.

📄 PDF Abstract BibTeX arXiv:2501.00142

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Minimalist Visual Inertial Odometry

2026-05-19 · Francesco Pasti, Jeremy Klotz, Nicola Bellotto, Shree K. Nayar arxiv

Visual-Inertial Odometry (VIO), which is critical to mobile robot navigation, uses cameras with a large number of pixels. Capturing and processing camera images requires significant resources. This work presents a minima…

Robot Navigation

Weakly-Supervised Open-Retrieval Conversational Question Answering

2021-03-03 · Chen Qu, Liu Yang, Cen Chen, W. Bruce Croft 외

Recent studies on Question Answering (QA) and Conversational QA (ConvQA) emphasize the role of retrieval: a system first retrieves evidence from a large collection and then extracts answers. This open-retrieval ConvQA se…

Conversational Question AnsweringQuestion AnsweringRetrieval

Adaptive remanufacturing for freeform surface parts based on linear laser scanner and robotic laser cladding

2024-08-10 · Robotics and Computer-Integrated Manufacturing 2024 8 · Wei Ma, TianliangHu, Chengrui Zhang, Qizhi Chen

Freeform surface parts play a significant role in the aerospace industry, the mold- manufacturing industry and the automobile industry, and it is energy-saving, material-saving, time-saving and environmentally beneficial…

Enhancing Video Summarization via Vision-Language Embedding

2017-07-01 · CVPR 2017 7 · Bryan A. Plummer, Matthew Brown, Svetlana Lazebnik

This paper addresses video summarization, or the problem of distilling a raw video into a shorter form while still capturing the original story. We show that visual representations supervised by freeform language make a …

Video Summarization

Minimalist Softmax Attention Provably Learns Constrained Boolean Functions

2025-05-26 · Jerry Yao-Chieh Hu, Xiwen Zhang, Maojiang Su, Zhao Song 외

We study the computational limits of learning $k$-bit Boolean functions (specifically, $\mathrm{AND}$, $\mathrm{OR}$, and their noisy variants), using a minimalist single-head softmax-attention mechanism, where $k=\Theta…