paper-with-me

Papers

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing

2026-07-30 · Jiajia Lin, Mingxuan Du, Tuowen Zhou, Benfeng Xu, Hongtao Xie arxiv

Text-to-image and personalized editing models now synthesize high-fidelity single-subject images with ease. Yet placing multiple named people into shared contact actions such as embrace, carry, or grapple still exposes major failures: fused limbs, invented extremities, and interpenetrating bodies. Existing evaluations largely overlook these anatomical and geometric issues, and VLM-as-a-judge checklists often saturate on Interaction while the errors remain obvious to humans. We introduce MPIE-Bench, a 2,500-sample benchmark of video-mined editing triplets spanning 405 scenes, 14 interaction categories, and four contact densities (C0-C3). We also propose MPIE-Eval, whose two new axes score contact-time geometry from a frozen public multi-person mesh reconstruction. Anatomy asks whether every human-like mass is explained by a complete set of reconstructed bodies, and Interaction asks whether the penetration and surface distance between those bodies match the contact the instruction asked for. Across ten editors, mesh Anatomy tops out at 0.65 and mesh Interaction at 0.72 on two different models, so no single editor is strong on both, while VLM checklists rate the same images above 0.95. A five-rater study confirms that both axes track human judgement more closely than a zero-shot VLM judge, and the rankings hold under ablation of every weight and threshold.

📄 PDF Abstract BibTeX arXiv:2607.27616

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Cardiac MRI Segmentation with Strong Anatomical Guarantees

2019-07-05 · Nathan Painchaud, Youssef Skandarani, Thierry Judge, Olivier Bernard 외

Recent publications have shown that the segmentation accuracy of modern-day convolutional neural networks (CNN) applied on cardiac MRI can reach the inter-expert variability, a great achievement in this area of research.…

Cardiac SegmentationMRI segmentationSegmentationvalid

Tractogram filtering of anatomically non-plausible fibers with geometric deep learning

2020-03-24 · Pietro Astolfi, Ruben Verhagen, Laurent Petit, Emanuele Olivetti 외

Tractograms are virtual representations of the white matter fibers of the brain. They are of primary interest for tasks like presurgical planning, and investigation of neuroplasticity or brain disorders. Each tractogram …

AnatomyDeep Learning

Context-aware virtual adversarial training for anatomically-plausible segmentation

2021-07-12 · Ping Wang, Jizong Peng, Marco Pedersoli, Yuanfeng Zhou 외

Despite their outstanding accuracy, semi-supervised segmentation methods based on deep neural networks can still yield predictions that are considered anatomically impossible by clinicians, for instance, containing holes…

Segmentation

Benchmarking Deep Learning-Based Reconstruction Methods for Photoacoustic Computed Tomography with Clinically Relevant Synthetic Datasets

2026-01-23 · Panpan Chen, Seonyeong Park, Gangwon Jeong, Refik Mert Cam 외 arxiv

Deep learning (DL)-based image reconstruction methods for photoacoustic computed tomography (PACT) have developed rapidly in recent years. However, most existing methods have not employed standardized datasets, and their…

Image Reconstruction

HumanScore: Benchmarking Human Motions in Generated Videos

2026-04-22 · Yusu Fang, Tiange Xiang, Tian Tan, Narayan Schuetz 외 arxiv

Recent advances in model architectures, compute, and data scale have driven rapid progress in video generation, producing increasingly realistic content. Yet, no prior method systematically measures how faithfully these …

Video Generation