Irrelevant Pixels are Everywhere: Find and Exclude Them for More Efficient Computer Vision
Computer vision is often performed using Convolutional Neural Networks (CNNs). CNNs are compute-intensive and challenging to deploy on power-contrained systems such as mobile and Internet-of-Things (IoT) devices. CNNs are compute-intensive because they indiscriminately compute many features on all pixels of the input image. We observe that, given a computer vision task, images often contain pixels that are irrelevant to the task. For example, if the task is looking for cars, pixels in the sky are not very useful. Therefore, we propose that a CNN be modified to only operate on relevant pixels to save computation and energy. We propose a method to study three popular computer vision datasets, finding that 48% of pixels are irrelevant. We also propose the focused convolution to modify a CNN's convolutional layers to reject the pixels that are marked irrelevant. On an embedded device, we observe no loss in accuracy, while inference latency, energy consumption, and multiply-add count are all reduced by about 45%.
Code (0)
등록된 구현이 없습니다.
Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Is happiness u-shaped in age everywhere? A methodological reconsideration for Europe
A recent contribution to research on age and well-being (Blanchflower 2021) found that the impact of age on happiness is "u-shaped" virtually everywhere: happiness declines towards middle age and subsequently rises, in a…
The Death of Schema Linking? Text-to-SQL in the Age of Well-Reasoned Language Models
Schema linking is a crucial step in Text-to-SQL pipelines. Its goal is to retrieve the relevant tables and columns of a target database for a user's query while disregarding irrelevant ones. However, imperfect schema lin…
Natural Language QueriesText to SQLText-To-SQLSuperpixel Segmentation using Dynamic and Iterative Spanning Forest
As constituent parts of image objects, superpixels can improve several higher-level operations. However, image segmentation methods might have their accuracy seriously compromised for reduced numbers of superpixels. We h…
ARCImage SegmentationSemantic SegmentationSuperpixelsOCM3D: Object-Centric Monocular 3D Object Detection
Image-only and pseudo-LiDAR representations are commonly used for monocular 3D object detection. However, methods based on them have shortcomings of either not well capturing the spatial relationships in neighbored image…
3D Object DetectionMonocular 3D Object DetectionObjectobject-detection+1RANSAC based three points algorithm for ellipse fitting of spherical object's projection
As the spherical object can be seen everywhere, we should extract the ellipse image accurately and fit it by implicit algebraic curve in order to finish the 3D reconstruction. In this paper, we propose a new ellipse fitt…
3D ReconstructionObject