paper-with-me

홈 › Papers

An Empirical Evaluation of Four Off-the-Shelf Proprietary Visual-Inertial Odometry Systems

2022-07-14 · Jungha Kim, Minkyeong Song, Yeoeun Lee, Moonkyeong Jung, Pyojin Kim

Commercial visual-inertial odometry (VIO) systems have been gaining attention as cost-effective, off-the-shelf six degrees of freedom (6-DoF) ego-motion tracking methods for estimating accurate and consistent camera pose data, in addition to their ability to operate without external localization from motion capture or global positioning systems. It is unclear from existing results, however, which commercial VIO platforms are the most stable, consistent, and accurate in terms of state estimation for indoor and outdoor robotic applications. We assess four popular proprietary VIO systems (Apple ARKit, Google ARCore, Intel RealSense T265, and Stereolabs ZED 2) through a series of both indoor and outdoor experiments where we show their positioning stability, consistency, and accuracy. We present our complete results as a benchmark comparison for the research community.

📄 PDF Abstract BibTeX arXiv:2207.06780

Code (0)

등록된 구현이 없습니다.

Tasks

State Estimation

Similar Papers 제목 키워드 기반

GenArena: How Can We Achieve Human-Aligned Evaluation for Visual Generation Tasks?

2026-02-05 · Ruihang Li, Leigang Qu, Jingxu Zhang, Dongnan Gui 외 arxiv

The rapid advancement of visual generation models has outpaced traditional evaluation approaches, necessitating the adoption of Vision-Language Models as surrogate judges. In this work, we systematically investigate the …

HARBOR: Holistic Adaptive Risk assessment model for BehaviORal healthcare

2025-12-21 · Aditya Siddhant arxiv

Behavioral healthcare risk assessment remains a challenging problem due to the highly multimodal nature of patient data and the temporal dynamics of mood and affective disorders. While large language models (LLMs) have d…

JMed48k: A Multi-Profession Japanese Medical Licensing Benchmark for Vision-Language Model Evaluation

2026-05-21 · Yue Xun, Junyu Liu, Qian Niu, Xinyi Wang 외 arxiv

We introduce JMed48k, a multi-profession Japanese healthcare licensing benchmark for evaluating vision-language models. Built from official PDF materials released by the Japanese Ministry of Health, Labour and Welfare, J…

Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains

2026-08-10 · Diandian Zhang, Tingyu Song, Lin Fu, Zheyuan Yang 외 hf

We introduce Sci-VBench, a comprehensive benchmark for evaluating knowledge- and reasoning-intensive video generation across scientific domains. It contains 1,253 expert-annotated examples spanning 60 subjects across fou…

Video Generation

VQ-VA World: Towards High-Quality Visual Question-Visual Answering

2025-11-25 · Chenhui Gou, Zilong Chen, Zeyu Wang, Feng Li 외 arxiv

This paper studies Visual Question-Visual Answering (VQ-VA): generating an image, rather than text, in response to a visual question -- an ability that has recently emerged in proprietary systems such as NanoBanana and G…