paper-with-me

Papers

Feature Tracks are not Zero-Mean Gaussian

2023-03-25 · Stephanie Tsuei, Wenjie Mo, Stefano Soatto

In state estimation algorithms that use feature tracks as input, it is customary to assume that the errors in feature track positions are zero-mean Gaussian. Using a combination of calibrated camera intrinsics, ground-truth camera pose, and depth images, it is possible to compute ground-truth positions for feature tracks extracted using an image processing algorithm. We find that feature track errors are not zero-mean Gaussian and that the distribution of errors is conditional on the type of motion, the speed of motion, and the image processing algorithm used to extract the tracks.

📄 PDF Abstract BibTeX arXiv:2303.14315

Code (0)

등록된 구현이 없습니다.

Tasks

State Estimation

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Performance-Guided Refinement for Visual Aerial Navigation using Editable Gaussian Splatting in FalconGym 2.0

2025-10-02 · Yan Miao, Ege Yuceel, Georgios Fainekos, Bardh Hoxha 외 arxiv

Visual policy design is crucial for aerial navigation. However, state-of-the-art visual policies often overfit to a single track and their performance degrades when track geometry changes. We develop FalconGym 2.0, a pho…

Scaling NVIDIA's Multi-speaker Multi-lingual TTS Systems with Zero-Shot TTS to Indic Languages

2024-01-24 · Akshit Arora, Rohan Badlani, Sungwon Kim, Rafael Valle 외

In this paper, we describe the TTS models developed by NVIDIA for the MMITS-VC (Multi-speaker, Multi-lingual Indic TTS with Voice Cloning) 2024 Challenge. In Tracks 1 and 2, we utilize RAD-MMM to perform few-shot TTS by …

Voice Cloning

Probabilistic Elastic Matching for Pose Variant Face Verification

2013-06-01 · CVPR 2013 6 · Haoxiang Li, Gang Hua, Zhe Lin, Jonathan Brandt 외

Pose variation remains to be a major challenge for realworld face recognition. We approach this problem through a probabilistic elastic matching method. We take a part based representation by extracting local features (e…

Face RecognitionFace Verification

The VoiceMOS Challenge 2023: Zero-shot Subjective Speech Quality Prediction for Multiple Domains

2023-10-04 · Erica Cooper, Wen-Chin Huang, Yu Tsao, Hsin-Min Wang 외

We present the second edition of the VoiceMOS Challenge, a scientific event that aims to promote the study of automatic prediction of the mean opinion score (MOS) of synthesized and processed speech. This year, we emphas…

Speech Synthesistext-to-speechText to SpeechText-To-Speech Synthesis

Time-Uniform Confidence Spheres for Means of Random Vectors

2023-11-14 · Ben Chugg, Hongjian Wang, Aaditya Ramdas

We study sequential mean estimation in $\mathbb{R}^d$. In particular, we derive time-uniform confidence spheres -- confidence sphere sequences (CSSs) -- which contain the mean of random vectors with high probability simu…

valid