paper-with-me

Papers

Watch your Up-Convolution: CNN Based Generative Deep Neural Networks are Failing to Reproduce Spectral Distributions

2020-03-03 · CVPR 2020 6 · Ricard Durall, Margret Keuper, Janis Keuper

Generative convolutional deep neural networks, e.g. popular GAN architectures, are relying on convolution based up-sampling methods to produce non-scalar outputs like images or video sequences. In this paper, we show that common up-sampling methods, i.e. known as up-convolution or transposed convolution, are causing the inability of such models to reproduce spectral distributions of natural training data correctly. This effect is independent of the underlying architecture and we show that it can be used to easily detect generated data like deepfakes with up to 100% accuracy on public benchmarks. To overcome this drawback of current generative models, we propose to add a novel spectral regularization term to the training optimization objective. We show that this approach not only allows to train spectral consistent GANs that are avoiding high frequency errors. Also, we show that a correct approximation of the frequency spectrum has positive effects on the training stability and output quality of generative networks.

📄 PDF Abstract BibTeX arXiv:2003.01826

Code (2)

cc-hpc-itwm/UpConv 공식 구현 pytorch
cc-hpc-itwm/DeepFakeDetection

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Dogecoin Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Watch Your Mouth: Silent Speech Recognition with Depth Sensing

2024-05-11 · Proceedings of the CHI Conference on Human Factors in Computing Systems 2024 5 · Xue Wang, Zixiong Su, Jun Rekimoto, Yang Zhang

Silent speech recognition is a promising technology that decodes human speech without requiring audio signals, enabling private human-computer interactions. In this paper, we propose Watch Your Mouth, a novel method that…

Deep LearningLipreadingSilent Speech Recognitionspeech-recognition+2

Watch your neighbors: Training statistically accurate chaotic systems with local phase space information

2026-05-14 · Joon-Hyuk Ko, Andrus Giraldo, Deok-Sun Lee arxiv

Chaotic systems pose fundamental challenges for data-driven dynamics discovery, as small modeling errors lead to exponentially growing trajectory discrepancies. Since exact long-term prediction is unattainable, it is nat…

Watch Your Steps: Local Image and Scene Editing by Text Instructions

2023-08-17 · Ashkan Mirzaei, Tristan Aumentado-Armstrong, Marcus A. Brubaker, Jonathan Kelly 외

Denoising diffusion models have enabled high-quality image generation and editing. We present a method to localize the desired edit region implicit in a text instruction. We leverage InstructPix2Pix (IP2P) and identify t…

DenoisingImage GenerationNeRF

What's on Your Plate? Inferring Chinese Cuisine Intake from Wearable IMUs

2025-11-07 · Jiaxi Yin, Pengcheng Wang, Han Ding, Fei Wang arxiv

Accurate food intake detection is vital for dietary monitoring and chronic disease prevention. Traditional self-report methods are prone to recall bias, while camera-based approaches raise concerns about privacy. Further…

Leveraging Watch-time Feedback for Short-Video Recommendations: A Causal Labeling Framework

2023-06-30 · Yang Zhang, Yimeng Bai, Jianxin Chang, Xiaoxue Zang 외

With the proliferation of short video applications, the significance of short video recommendations has vastly increased. Unlike other recommendation scenarios, short video recommendation systems heavily rely on feedback…

Recommendation Systems