paper-with-me

홈 › Papers

HandEval: Taking the First Step Towards Hand Quality Evaluation in Generated Images

2025-10-10 · Zichuan Wang, Bo Peng, Songlin Yang, Zhenchen Tang, Jing Dong arxiv

Although recent text-to-image (T2I) models have significantly improved the overall visual quality of generated images, they still struggle in the generation of accurate details in complex local regions, especially human hands. Generated hands often exhibit structural distortions and unrealistic textures, which can be very noticeable even when the rest of the body is well-generated. However, the quality assessment of hand regions remains largely neglected, limiting downstream task performance like human-centric generation quality optimization and AIGC detection. To address this, we propose the first quality assessment task targeting generated hand regions and showcase its abundant downstream applications. We first introduce the HandPair dataset for training hand quality assessment models. It consists of 48k images formed by high- and low-quality hand pairs, enabling low-cost, efficient supervision without manual annotation. Based on it, we develop HandEval, a carefully designed hand-specific quality assessment model. It leverages the powerful visual understanding capability of Multimodal Large Language Model (MLLM) and incorporates prior knowledge of hand keypoints, gaining strong perception of hand quality. We further construct a human-annotated test set with hand images from various state-of-the-art (SOTA) T2I models to validate its quality evaluation capability. Results show that HandEval aligns better with human judgments than existing SOTA methods. Furthermore, we integrate HandEval into image generation and AIGC detection pipelines, prominently enhancing generated hand realism and detection accuracy, respectively, confirming its universal effectiveness in downstream applications. Code and dataset will be available.

📄 PDF Abstract BibTeX arXiv:2510.08978

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

Progressive Distillation for Fast Sampling of Diffusion Models

2022-02-01 · ICLR 2022 4 · Tim Salimans, Jonathan Ho

Diffusion models have recently shown great promise for generative modeling, outperforming GANs on perceptual quality and autoregressive models at density estimation. A remaining downside is their slow sampling time: gene…

Density EstimationImage Generation

Uncalibrated Deflectometry with a Mobile Device on Extended Specular Surfaces

2019-07-24 · Florian Willomitzer, Chia-Kai Yeh, Vikas Gupta, William Spies 외

We introduce a system and methods for the three-dimensional measurement of extended specular surfaces with high surface normal variations. Our system consists only of a mobile hand held device and exploits screen and fro…

Background Matting: The World is Your Green Screen

2020-04-01 · CVPR 2020 6 · Soumyadip Sengupta, Vivek Jayaram, Brian Curless, Steve Seitz 외

We propose a method for creating a matte -- the per-pixel foreground color and alpha -- of a person by taking photos or videos in an everyday setting with a handheld camera. Most existing matting methods require a green …

Image Matting

FedDAG: Federated DAG Structure Learning

2021-12-07 · Erdun Gao, Junjia Chen, Li Shen, Tongliang Liu 외

To date, most directed acyclic graphs (DAGs) structure learning approaches require data to be stored in a central server. However, due to the consideration of privacy protection, data owners gradually refuse to share the…

Causal Discovery

SynManDex: Synthesizing Human-like Dexterous Grasps from Synthetic Human Pre-Grasps

2026-06-08 · Yanming Shao, Zanxin Chen, Wenwei Lin, Mingjie Zhou 외 arxiv

Human hand-object interactions encode functional intent, but direct transfer to robotic hands often fails under morphology, contact, and reachability constraints. We present SynManDex, a synthetic pipeline that uses gene…