Dual-Representation Interaction Driven Image Quality Assessment with Restoration Assistance
No-Reference Image Quality Assessment for distorted images has always been a challenging problem due to image content variance and distortion diversity. Previous IQA models mostly encode explicit single-quality features of synthetic images to obtain quality-aware representations for quality score prediction. However, performance decreases when facing real-world distortion and restored images from restoration models. The reason is that they do not consider the degradation factors of the low-quality images adequately. To address this issue, we first introduce the DRI method to obtain degradation vectors and quality vectors of images, which separately model the degradation and quality information of low-quality images. After that, we add the restoration network to provide the MOS score predictor with degradation information. Then, we design the Representation-based Semantic Loss (RS Loss) to assist in enhancing effective interaction between representations. Extensive experimental results demonstrate that the proposed method performs favorably against existing state-of-the-art models on both synthetic and real-world datasets.
Code (1)
Tasks
DiversityImage Quality AssessmentNo-Reference Image Quality AssessmentSimilar Papers 제목 키워드 기반
Learning a Unified Degradation-aware Representation Model for Multi-modal Image Fusion
All-in-One Degradation-Aware Fusion Models (ADFMs), a class of multi-modal image fusion models, address complex scenes by mitigating degradations from source images and generating high-quality fused images. Mainstream AD…
Infrared And Visible Image FusionDualNeRF: Text-Driven 3D Scene Editing via Dual-Field Representation
Recently, denoising diffusion models have achieved promising results in 2D image generation and editing. Instruct-NeRF2NeRF (IN2N) introduces the success of diffusion into 3D scene editing through an "Iterative dataset u…
3D scene EditingDenoisingImage GenerationNeRFDual Audio-Centric Modality Coupling for Talking Head Generation
The generation of audio-driven talking head videos is a key challenge in computer vision and graphics, with applications in virtual avatars and digital media. Traditional approaches often struggle with capturing the comp…
NeRFTalking Head Generationtext-to-speechText to SpeechDual Diffusion Models for Multi-modal Guided 3D Avatar Generation
Generating high-fidelity 3D avatars from text or image prompts is highly sought after in virtual reality and human-computer interaction. However, existing text-driven methods often rely on iterative Score Distillation Sa…
Computational EfficiencyPVLR: Prompt-driven Visual-Linguistic Representation Learning for Multi-Label Image Recognition
Multi-label image recognition is a fundamental task in computer vision. Recently, vision-language models have made notable advancements in this area. However, previous methods often failed to effectively leverage the ric…
Multi-Label Image RecognitionRepresentation Learning