paper-with-me

Papers

HYPE-C: Evaluating Image Completion Models Through Standardized Crowdsourcing

2021-01-01 · Emily Walters, Weifeng Chen, Jia Deng

A significant obstacle to the development of new image completion models is the lack of a standardized evaluation metric that reflects human judgement. Recent work has proposed the use of human evaluation for image synthesis models, allowing for a reliable method to evaluate the visual quality of generated images. However, there does not yet exist a standardized human evaluation protocol for image completion. In this work, we propose such a protocol. We also provide experimental results of our evaluation method applied to many of the current state-of-the-art generative image models and compare these results to various automated metrics. Our evaluation yields a number of interesting findings. Notably, GAN-based image completion models are outperformed by autoregressive approaches.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

A Standardized Framework For Evaluating Gene Expression Generative Models

2026-03-11 · Andrea Rubbi, Andrea Giuseppe Di Francesco, Mohammad Lotfollahi, Pietro Liò arxiv

The rapid development of generative models for single-cell gene expression data has created an urgent need for standardised evaluation frameworks. Current evaluation practices suffer from inconsistent metric implementati…

SQBench: A Benchmark for Evaluating Task Delivery by Language-Model Agents in Production-Oriented Workflows

2026-07-25 · Summer Sun arxiv

Existing evaluations of large language models cover knowledge, reasoning, coding, and tool use, but they rarely treat a verifiable deliverable produced within a constrained workflow as the unit of evaluation. We introduc…

DeOcc-1-to-3: 3D De-Occlusion from a Single Image via Self-Supervised Multi-View Diffusion

2025-06-26 · Yansong Qu, Shaohui Dai, Xinyang Li, Yuze Wang 외

Reconstructing 3D objects from a single image is a long-standing challenge, especially under real-world occlusions. While recent diffusion-based view synthesis models can generate consistent novel views from a single RGB…

3D Reconstruction

UnCLe: Unsupervised Continual Learning of Depth Completion

2024-10-23 · Suchisrit Gangopadhyay, Xien Chen, Michael Chu, Patrick Rim 외

We propose UnCLe, a standardized benchmark for Unsupervised Continual Learning of a multimodal depth estimation task: Depth completion aims to infer a dense depth map from a pair of synchronized RGB image and sparse dept…

Continual LearningDepth CompletionDepth Estimation

End-to-End Evaluation and Governance of an EHR-Embedded AI Agent for Clinicians

2026-04-30 · Aaryan Shah, Andrew Hines, Alexia Downs, Denis Bajet 외 arxiv

Clinical AI systems require not just point-in-time evaluation but continuous governance: the ongoing practice of monitoring, evaluating, iterating, and re-evaluating performance throughout deployment. We present an end-t…