paper-with-me

Papers

A Hybrid Model-Based and Model-Free Framework for Active Multi-View Viewpoint Optimization in Sonar Target Recognition

2026-06-13 · Yongkyoon Park, Jane Shin arxiv

This paper presents a hybrid model-based and model-free framework for active multi-view target recognition using forward-looking sonar. A convolutional neural network (CNN) provides data-driven observation likelihoods, while Radon-based orientation estimation enables viewpoint-aware sensing without requiring angle annotations. During training, an information-gain-based reward guides a Proximal Policy Optimization (PPO) agent to learn a belief-aware viewpoint selection policy offline. At deployment, the learned policy performs real-time viewpoint selection using only CNN-based belief updates, eliminating the need for computationally expensive online POMDP tree search. Experiments on a marine-debris forward-looking sonar dataset demonstrate that the proposed approach achieves competitive recognition accuracy while reducing sensing steps and motion cost compared to model-based baselines.

📄 PDF Abstract BibTeX arXiv:2606.15373

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

iButter: Neural Interactive Bullet Time Generator for Human Free-viewpoint Rendering

2021-08-12 · Liao Wang, Ziyu Wang, Pei Lin, Yuheng Jiang 외

Generating ``bullet-time'' effects of human free-viewpoint videos is critical for immersive visual effects and VR/AR experience. Recent neural advances still lack the controllable and interactive bullet-time design abili…

NeRFVideo Generation

If I Hear You Correctly: Building and Evaluating Interview Chatbots with Active Listening Skills

2020-02-05 · Ziang Xiao, Michelle X. Zhou, Wenxi Chen, Huahai Yang 외

Interview chatbots engage users in a text-based conversation to draw out their views and opinions. It is, however, challenging to build effective interview chatbots that can handle user free-text responses to open-ended …

Chatbot

Active3D: Active High-Fidelity 3D Reconstruction via Hierarchical Uncertainty Quantification

2025-11-25 · Yan Li, Yingzhao Li, Gim Hee Lee arxiv

In this paper, we present an active exploration framework for high-fidelity 3D reconstruction that incrementally builds a multi-level uncertainty space and selects next-best-views through an uncertainty-driven motion pla…

3D Reconstruction

Multi-Dimensional Reconfigurable, Physically Composable Hybrid Diffractive Optical Neural Network

2024-11-08 · Ziang Yin, Yu Yao, Jeff Zhang, Jiaqi Gu

Diffractive optical neural networks (DONNs), leveraging free-space light wave propagation for ultra-parallel, high-efficiency computing, have emerged as promising artificial intelligence (AI) accelerators. However, their…

DeepScan: A Training-Free Framework for Visually Grounded Reasoning in Large Vision-Language Models

2026-03-04 · Yangfu Li, Hongjian Zhan, Jiawei Chen, Yuning Gong 외 arxiv

Humans can robustly localize visual evidence and provide grounded answers even in noisy environments by identifying critical cues and then relating them to the full context in a bottom-up manner. Inspired by this, we pro…