paper-with-me

홈 › Papers

HybVIO: Pushing the Limits of Real-time Visual-inertial Odometry

2021-06-22 · Otto Seiskari, Pekka Rantalankila, Juho Kannala, Jerry Ylilammi, Esa Rahtu, Arno Solin

We present HybVIO, a novel hybrid approach for combining filtering-based visual-inertial odometry (VIO) with optimization-based SLAM. The core of our method is highly robust, independent VIO with improved IMU bias modeling, outlier rejection, stationarity detection, and feature track selection, which is adjustable to run on embedded hardware. Long-term consistency is achieved with a loosely-coupled SLAM module. In academic benchmarks, our solution yields excellent performance in all categories, especially in the real-time use case, where we outperform the current state-of-the-art. We also demonstrate the feasibility of VIO for vehicular tracking on consumer-grade hardware using a custom dataset, and show good performance in comparison to current commercial VISLAM alternatives. An open-source implementation of the HybVIO method is available at https://github.com/SpectacularAI/HybVIO

📄 PDF Abstract BibTeX arXiv:2106.11857

Code (1)

SpectacularAI/HybVIO 공식 구현

Similar Papers 제목 키워드 기반

ReViCo: Unveiling the Limitations of VLMs in Visual Text Understanding via Error Correction

2026-08-27 · Bojun Zhang, Junhong Liang, Feifei Zhai, Fengxian Ji 외 arxiv

Vision Language Models (VLMs) have shown great success in general visual tasks, yet they still struggle to deeply understand text within images. In this paper, we introduce ReViCo (Real Visual Correction), a benchmark de…

Conceptual 12M: Pushing Web-Scale Image-Text Pre-Training To Recognize Long-Tail Visual Concepts

2021-02-17 · CVPR 2021 1 · Soravit Changpinyo, Piyush Sharma, Nan Ding, Radu Soricut

The availability of large-scale image captioning and visual question answering datasets has contributed significantly to recent successes in vision-and-language pre-training. However, these datasets are often collected w…

Caption GenerationDiversityImage CaptioningQuestion Answering+2

Pushing the Limits of Radiology with Joint Modeling of Visual and Textual Information

2018-07-01 · ACL 2018 7 · Sonit Singh

Recently, there has been increasing interest in the intersection of computer vision and natural language processing. Researchers have studied several interesting tasks, including generating text descriptions from images …

Image ClassificationMachine TranslationObject DetectionQuestion Answering+5

Omnipush: accurate, diverse, real-world dataset of pushing dynamics with RGB-D video

2019-10-01 · Maria Bauza, Ferran Alet, Yen-Chen Lin, Tomas Lozano-Perez 외

Pushing is a fundamental robotic skill. Existing work has shown how to exploit models of pushing to achieve a variety of tasks, including grasping under uncertainty, in-hand manipulation and clearing clutter. Such models…

Meta-LearningVideo Prediction

Pushing the Limits of Learning-based Traversability Analysis for Autonomous Driving on CPU

2022-06-07 · Daniel Fusaro, Emilio Olivastri, Daniele Evangelista, Marco Imperoli 외

Self-driving vehicles and autonomous ground robots require a reliable and accurate method to analyze the traversability of the surrounding environment for safe navigation. This paper proposes and evaluates a real-time ma…

Autonomous DrivingCPU