paper-with-me

홈 › Papers

MRSCAtt: A Spatio-Channel Attention-Guided Network for Mars Rover Image Classification

2021-06-12 · Anirudh S Chakravarthy, Roshan Roy, Praveen Ravirathinam

As the exploration of human beings pushes deeper into the galaxy, the classification of images from space and other planets is becoming an increasingly critical task. Image classification on these planetary images can be very challenging due to differences in hue, quality, illumination, and clarity when compared to images captured on Earth. In this work, we try to bridge this gap by developing a deep learning network, MRSCAtt (Mars Rover Spatial and Channel Attention), which jointly uses spatial and channel attention to accurately classify images. We use images taken by NASA's Curiosity rover on Mars as a dataset to show the superiority of our approach by achieving state-of-the-art results with 81.53% test set accuracy on the MSL Surface Dataset, outperforming other methods. To necessitate the use of spatial and channel attention, we perform an ablation study to show the effectiveness of each of the components. We further show robustness of our approach by validating with images taken aboard NASA's recently-landed Perseverance rover.

📄 PDF Abstract BibTeX

Code (1)

anirudh-chakravarthy/MRSCAtt pytorch

Tasks

image-classificationImage Classification

Similar Papers 제목 키워드 기반

Beyond Pedestrians: Caption-Guided CLIP Framework for High-Difficulty Video-based Person Re-Identification

2026-04-09 · Shogo Hamano, Shunya Wakasugi, Tatsuhito Sato, Sayaka Nakamura arxiv

In recent years, video-based person Re-Identification (ReID) has gained attention for its ability to leverage spatiotemporal cues to match individuals across non-overlapping cameras. However, current methods struggle wit…

Person Re-Identification

KeyRe-ID: Keypoint-Guided Person Re-Identification using Part-Aware Representation in Videos

2025-07-10 · Jinseong Kim, Jeonghoon Song, Gyeongseon Baek, Byeongjoon Noh

We propose \textbf{KeyRe-ID}, a keypoint-guided video-based person re-identification framework consisting of global and local branches that leverage human keypoints for enhanced spatiotemporal representation learning. Th…

Person Re-IdentificationRepresentation LearningVideo-Based Person Re-Identification

MARS: Mask Attention Refinement with Sequential Quadtree Nodes for Car Damage Instance Segmentation

2023-05-01 · Teerapong Panboonyuen, Naphat Nithisopa, Panin Pienroj, Laphonchai Jirachuphun 외

Evaluating car damages from misfortune is critical to the car insurance industry. However, the accuracy is still insufficient for real-world applications since the deep learning network is not designed for car damage ima…

Instance SegmentationSegmentationSemantic Segmentation

VidLaDA: Bidirectional Diffusion Large Language Models for Efficient Video Understanding

2026-01-25 · Zhihao He, Tieyuan Chen, Kangyu Wang, Ziran Qin 외 arxiv

Current Video Large Language Models (Video LLMs) typically encode frames via a vision encoder and employ an autoregressive (AR) LLM for understanding and generation. However, this AR paradigm inevitably faces a dual effi…

MARs: Multi-view Attention Regularizations for Patch-based Feature Recognition of Space Terrain

2024-10-07 · Timothy Chase Jr, Karthik Dantu

The visual detection and tracking of surface terrain is required for spacecraft to safely land on or navigate within close proximity to celestial objects. Current approaches rely on template matching with pre-gathered pa…

AttributeMetric LearningNavigateTemplate Matching