paper-with-me

Papers

EyeSim-VQA: A Free-Energy-Guided Eye Simulation Framework for Video Quality Assessment

2025-06-13 · Zhaoyang Wang, Wen Lu, Jie Li, Lihuo He, Maoguo Gong, Xinbo Gao

Free-energy-guided self-repair mechanisms have shown promising results in image quality assessment (IQA), but remain under-explored in video quality assessment (VQA), where temporal dynamics and model constraints pose unique challenges. Unlike static images, video content exhibits richer spatiotemporal complexity, making perceptual restoration more difficult. Moreover, VQA systems often rely on pre-trained backbones, which limits the direct integration of enhancement modules without affecting model stability. To address these issues, we propose EyeSimVQA, a novel VQA framework that incorporates free-energy-based self-repair. It adopts a dual-branch architecture, with an aesthetic branch for global perceptual evaluation and a technical branch for fine-grained structural and semantic analysis. Each branch integrates specialized enhancement modules tailored to distinct visual inputs-resized full-frame images and patch-based fragments-to simulate adaptive repair behaviors. We also explore a principled strategy for incorporating high-level visual features without disrupting the original backbone. In addition, we design a biologically inspired prediction head that models sweeping gaze dynamics to better fuse global and local representations for quality prediction. Experiments on five public VQA benchmarks demonstrate that EyeSimVQA achieves competitive or superior performance compared to state-of-the-art methods, while offering improved interpretability through its biologically grounded design.

📄 PDF Abstract BibTeX arXiv:2506.11549

Code (0)

등록된 구현이 없습니다.

Tasks

Image Quality AssessmentVideo Quality AssessmentVisual Question Answering (VQA)

Similar Papers 제목 키워드 기반

Enhancing Reliability in LLM-Integrated Robotic Systems: A Unified Approach to Security and Safety

2025-09-02 · Wenxiao Zhang, Xiangrui Kong, Conan Dewitt, Thomas Bräunl 외 arxiv

Integrating large language models (LLMs) into robotic systems has revolutionised embodied artificial intelligence, enabling advanced decision-making and adaptability. However, ensuring reliability, encompassing both secu…

FEAT: Free energy Estimators with Adaptive Transport

2025-04-15 · Jiajun He, Yuanqi Du, Francisco Vargas, Yuanqing Wang 외

We present Free energy Estimators with Adaptive Transport (FEAT), a novel framework for free energy estimation -- a critical challenge across scientific domains. FEAT leverages learned transports implemented via stochast…

PathRIR: Physics-Guided Acoustic Path Selection and Late-Tail Compensation for Fast Room Impulse Response Simulation

2026-07-25 · Shaoheng Xu, Chunyi Sun, Jihui Zhang, Amy Bastine 외 arxiv

Image-source-method (ISM)-based room impulse response (RIR) simulation is a useful and physically interpretable tool for acoustic scene modeling, but full-order ISM becomes computationally expensive as the reflection ord…

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment

2026-01-29 · Xiuyu Li, Jinkai Zhang, Mingyang Yi, Yu Li 외 arxiv

Reinforcement Learning (RL) post-training alignment for language models is effective, but also costly and unstable in practice, owing to its complicated training process. To address this, we propose a training-free infer…

Reinforcement Learning

Graph Distance as Surprise: Free Energy Minimization in Knowledge Graph Reasoning

2025-12-01 · Gaganpreet Jhajj, Fuhua Lin arxiv

In this work, we propose that reasoning in knowledge graph (KG) networks can be guided by surprise minimization. Entities that are close in graph distance will have lower surprise than those farther apart. This connects …

Reinforcement Learning