paper-with-me

홈 › Papers

Geometry-based spherical JND modeling for 360$^\circ$ display

2023-03-07 · Hongan Wei, Jiaqi Liu, Bo Chen, Liqun Lin, Weiling Chen, Tiesong Zhao

360$^\circ$ videos have received widespread attention due to its realistic and immersive experiences for users. To date, how to accurately model the user perceptions on 360$^\circ$ display is still a challenging issue. In this paper, we exploit the visual characteristics of 360$^\circ$ projection and display and extend the popular just noticeable difference (JND) model to spherical JND (SJND). First, we propose a quantitative 2D-JND model by jointly considering spatial contrast sensitivity, luminance adaptation and texture masking effect. In particular, our model introduces an entropy-based region classification and utilizes different parameters for different types of regions for better modeling performance. Second, we extend our 2D-JND model to SJND by jointly exploiting latitude projection and field of view during 360$^\circ$ display. With this operation, SJND reflects both the characteristics of human vision system and the 360$^\circ$ display. Third, our SJND model is more consistent with user perceptions during subjective test and also shows more tolerance in distortions with fewer bit rates during 360$^\circ$ video compression. To further examine the effectiveness of our SJND model, we embed it in Versatile Video Coding (VVC) compression. Compared with the state-of-the-arts, our SJND-VVC framework significantly reduced the bit rate with negligible loss in visual quality.

📄 PDF Abstract BibTeX arXiv:2303.03703

Code (0)

등록된 구현이 없습니다.

Tasks

Video Compression

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

SGAT4PASS: Spherical Geometry-Aware Transformer for PAnoramic Semantic Segmentation

2023-06-06 · XueWei Li, Tao Wu, Zhongang Qi, Gaoang Wang 외

As an important and challenging problem in computer vision, PAnoramic Semantic Segmentation (PASS) gives complete scene perception based on an ultra-wide angle of view. Usually, prevalent PASS methods with 2D panoramic i…

Semantic Segmentation

Lightweight and Robust Multi-Channel End-to-End Speech Recognition with Spherical Harmonic Transform

2025-06-13 · Xiangzhu Kong, Huang Hao, Zhijian Ou

This paper presents SHTNet, a lightweight spherical harmonic transform (SHT) based framework, which is designed to address cross-array generalization challenges in multi-channel automatic speech recognition (ASR) through…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)channel selectionspeech-recognition+1

Pano-AVQA: Grounded Audio-Visual Question Answering on 360$^\circ$ Videos

2021-10-11 · Heeseung Yun, Youngjae Yu, Wonsuk Yang, Kangil Lee 외

360$^\circ$ videos convey holistic views for the surroundings of a scene. It provides audio-visual cues beyond pre-determined normal field of views and displays distinctive spatial relations on a sphere. However, previou…

Audio-visual Question AnsweringQuestion AnsweringRelationVisual Question Answering+1

PanoGRF: Generalizable Spherical Radiance Fields for Wide-baseline Panoramas

2023-06-02 · NeurIPS 2023 11 · Zheng Chen, Yan-Pei Cao, Yuan-Chen Guo, Chen Wang 외

Achieving an immersive experience enabling users to explore virtual environments with six degrees of freedom (6DoF) is essential for various applications such as virtual reality (VR). Wide-baseline panoramas are commonly…

Depth Estimation

3D Scene Geometry Estimation from 360$^\circ$ Imagery: A Survey

2024-01-17 · Thiago Lopes Trugillo da Silveira, Paulo Gamarra Lessa Pinto, Jeffri Erwin Murrugarra Llerena, Claudio Rosito Jung

This paper provides a comprehensive survey on pioneer and state-of-the-art 3D scene geometry estimation methodologies based on single, two, or multiple images captured under the omnidirectional optics. We first revisit t…

Simultaneous Localization and MappingStereo MatchingSurvey