paper-with-me

Papers

iPhoneBlur: A Difficulty-Stratified Benchmark for Consumer Device Motion Deblurring

2026-05-07 · Abdullah Al Shafi, Kazi Saeed Alam arxiv

Motion blur restoration on consumer mobile devices is typically evaluated using aggregate metrics that obscure performance variation across blur difficulty, masking model behavior under real deployment conditions. This work introduces iPhoneBlur, a difficulty-stratified benchmark of 7,400 image pairs synthesized from high-framerate iPhone 17 Pro videos captured in diverse real-world scenarios. Samples are partitioned into Easy, Medium, and Hard categories through PSNR-guided adaptive temporal windowing, with stratification validated by monotonic 2.2x increase in optical flow magnitude across tiers. Each sample includes comprehensive metadata enabling investigation of ISP-aware and difficulty-adaptive restoration strategies. Spectral analysis confirms synthesized blur exhibits high-frequency suppression patterns consistent with authentic motion degradation. Evaluation of six architectures reveals consistent 7-9 dB performance degradation from Easy to Hard subsets, a substantial gap entirely hidden by aggregate reporting. The benchmark further exposes a domain gap between professional and consumer cameras which targeted fine-tuning substantially recovers. By coupling difficulty stratification with deployment-critical metadata, iPhoneBlur enables systematic assessment of model reliability and failure modes for resource-constrained edge systems.

📄 PDF Abstract BibTeX arXiv:2605.05990

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AnalogSAGE: Self-evolving Analog Design Multi-Agents with Stratified Memory and Grounded Experience

2025-12-27 · Zining Wang, Jian Gao, Weimin Fu, Xiaolong Guo 외 arxiv

Analog circuit design remains a knowledge- and experience-intensive process that relies heavily on human intuition for topology generation and device parameter tuning. Existing LLM-based approaches typically depend on pr…

ConsumerBench: Benchmarking Generative AI Applications on End-User Devices

2025-06-21 · Yile Gu, Rohan Kadekodi, Hoang Nguyen, Keisuke Kamahori 외

The recent shift in Generative AI (GenAI) applications from cloud-only environments to end-user devices introduces new challenges in resource management, system efficiency, and user experience. This paper presents Consum…

BenchmarkingCPUGPUScheduling

Stratified Sampling for Extreme Multi-Label Data

2021-03-05 · Maximillian Merrillees, Lan Du

Extreme multi-label classification (XML) is becoming increasingly relevant in the era of big data. Yet, there is no method for effectively generating stratified partitions of XML datasets. Instead, researchers typically …

Extreme Multi-Label ClassificationMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION

Region-Specific Calibration Achieves Excellent Inter-Device Reliability for Smartphone Dermatology: A Multi-Device Benchmark on Korean Facial Skin

2025-12-26 · Sungwoo Kang, Jong-Kook Kim arxiv

Background: Smartphone-based dermatology requires inter-device colorimetric reliability that holds across calibration regimes, yet quantitative multi-device benchmarks remain scarce. Materials and Methods: We analyzed ma…

REDACT: A Systematically Controlled Multilingual Benchmark for Personal Information Detection

2026-06-18 · Guneesh Vats, Anubha Agrawal, Shikha Singhal, Ajita Dash 외 arxiv

Benchmark infrastructure for personally identifiable information (PII) detection remains limited: existing corpora cover few entity types, use ad hoc generation conditions, and do not show which surface conditions cause …