paper-with-me

Papers

Dual-Representation Interaction Driven Image Quality Assessment with Restoration Assistance

2024-11-26 · Jingtong Yue, Xin Lin, Zijiu Yang, Chao Ren

No-Reference Image Quality Assessment for distorted images has always been a challenging problem due to image content variance and distortion diversity. Previous IQA models mostly encode explicit single-quality features of synthetic images to obtain quality-aware representations for quality score prediction. However, performance decreases when facing real-world distortion and restored images from restoration models. The reason is that they do not consider the degradation factors of the low-quality images adequately. To address this issue, we first introduce the DRI method to obtain degradation vectors and quality vectors of images, which separately model the degradation and quality information of low-quality images. After that, we add the restoration network to provide the MOS score predictor with degradation information. Then, we design the Representation-based Semantic Loss (RS Loss) to assist in enhancing effective interaction between representations. Extensive experimental results demonstrate that the proposed method performs favorably against existing state-of-the-art models on both synthetic and real-world datasets.

📄 PDF Abstract BibTeX arXiv:2411.17390

Code (1)

Jingtong0527/DRI-IQA 공식 구현

Tasks

DiversityImage Quality AssessmentNo-Reference Image Quality Assessment

Similar Papers 제목 키워드 기반

Learning a Unified Degradation-aware Representation Model for Multi-modal Image Fusion

2025-03-10 · Haolong Ma, Hui Li, Chunyang Cheng, Zeyang Zhang 외

All-in-One Degradation-Aware Fusion Models (ADFMs), a class of multi-modal image fusion models, address complex scenes by mitigating degradations from source images and generating high-quality fused images. Mainstream AD…

Infrared And Visible Image Fusion

DualNeRF: Text-Driven 3D Scene Editing via Dual-Field Representation

2025-02-22 · Yuxuan Xiong, Yue Shi, Yishun Dou, Bingbing Ni

Recently, denoising diffusion models have achieved promising results in 2D image generation and editing. Instruct-NeRF2NeRF (IN2N) introduces the success of diffusion into 3D scene editing through an "Iterative dataset u…

3D scene EditingDenoisingImage GenerationNeRF

Dual Audio-Centric Modality Coupling for Talking Head Generation

2025-03-26 · Ao Fu, Ziqi Ni, Yi Zhou

The generation of audio-driven talking head videos is a key challenge in computer vision and graphics, with applications in virtual avatars and digital media. Traditional approaches often struggle with capturing the comp…

NeRFTalking Head Generationtext-to-speechText to Speech

Dual Diffusion Models for Multi-modal Guided 3D Avatar Generation

2026-03-04 · Hong Li, Yutang Feng, Minqi Meng, Yichen Yang 외 arxiv

Generating high-fidelity 3D avatars from text or image prompts is highly sought after in virtual reality and human-computer interaction. However, existing text-driven methods often rely on iterative Score Distillation Sa…

Computational Efficiency

PVLR: Prompt-driven Visual-Linguistic Representation Learning for Multi-Label Image Recognition

2024-01-31 · Hao Tan, Zichang Tan, Jun Li, Jun Wan 외

Multi-label image recognition is a fundamental task in computer vision. Recently, vision-language models have made notable advancements in this area. However, previous methods often failed to effectively leverage the ric…

Multi-Label Image RecognitionRepresentation Learning