paper-with-me

Papers

SynthForensics: Benchmarking and Evaluating People-Centric Synthetic Video Deepfakes

2026-02-04 · Roberto Leotta, Salvatore Alfio Sambataro, Claudio Vittorio Ragaglia, Mirko Casu, Yuri Petralia, Francesco Guarnera, Luca Guarnera, Sebastiano Battiato arxiv

Modern T2V/I2V generators synthesize people increasingly hard to distinguish from authentic footage, while current evaluation suites lag: legacy benchmarks target manipulation-based forgeries, and recent synthetic-video benchmarks prioritize scale over realistic human depiction. We introduce SynthForensics, a people-centric benchmark of $20{,}445$ videos from 8 T2V and 7 I2V open-source generators, paired-source from FF++/DFD reals, two-stage human-validated, in four compression versions with full metadata. In our paired-comparison human study, raters prefer SynthForensics in $71$--$77\%$ of head-to-head comparisons against each of nine existing synthetic-video benchmarks, while facial-quality metrics fall within the FF++/DFD baseline range. Across 15 detectors and three protocols, face-based methods drop $13$--$55$ AUC points (mean $27$) from FF++ to SynthForensics and a further $23$ under aggressive compression; fine-tuning closes the gap at a backward cost on legacy benchmarks; training from scratch shows synthetic and manipulation features largely disjoint for most detectors. We release dataset, pipeline, and code.

📄 PDF Abstract BibTeX arXiv:2602.04939

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PeopleSansPeople: A Synthetic Data Generator for Human-Centric Computer Vision

2021-12-17 · Salehe Erfanian Ebadi, You-Cyuan Jhang, Alex Zook, Saurav Dhakad 외

In recent years, person detection and human pose estimation have made great strides, helped by large-scale labeled datasets. However, these datasets had no guarantees or analysis of human activities, poses, or context di…

Human DetectionPose EstimationSemantic SegmentationTransfer Learning

Be More Specific: Evaluating Object-centric Realism in Synthetic Images

2025-01-01 · CVPR 2025 1 · Anqi Liang, Ciprian Corneanu, Qianli Feng, Giorgio Giannone 외

Evaluation of synthetic images is important for both model development and selection. An ideal evaluation should be specific, accurate and aligned with human perception. This paper addresses the problem of evaluating…

Object

MemoryCD: Benchmarking Long-Context User Memory of LLM Agents for Lifelong Cross-Domain Personalization

2026-03-26 · Weizhi Zhang, Xiaokai Wei, Wei-Chieh Huang, Zheng Hui 외 arxiv

Recent advancements in Large Language Models (LLMs) have expanded context windows to million-token scales, yet benchmarks for evaluating memory remain limited to short-session synthetic dialogues. We introduce \textsc{Me…

Scales++: Compute Efficient Evaluation Subset Selection with Cognitive Scales Embeddings

2025-10-30 · Andrew M. Bean, Nabeel Seedat, Shengzhuang Chen, Jonathan Richard Schwarz arxiv

The prohibitive cost of evaluating large language models (LLMs) on comprehensive benchmarks necessitates the creation of small yet representative data subsets (i.e., tiny benchmarks) that enable efficient assessment whil…

Benchmarking Synthetic Tabular Data: A Multi-Dimensional Evaluation Framework

2025-04-02 · Andrey Sidorenko, Michael Platzer, Mario Scriminaci, Paul Tiwald

Evaluating the quality of synthetic data remains a key challenge for ensuring privacy and utility in data-driven research. In this work, we present an evaluation framework that quantifies how well synthetic data replicat…

BenchmarkingSynthetic Data Generation