paper-with-me

홈 › Papers

From SRA to Self-Flow: Data Augmentation or Self-Supervision?

2026-07-02 · Dengyang Jiang, Mengmeng Wang, Harry Yang, Jingdong Wang arxiv

Representation alignment has become an effective way to accelerate diffusion transformer training and improve generation quality. Recent self-alignment methods, such as SRA and Self-Flow, further remove the dependency on external pretrained encoders by constructing alignment within the diffusion model itself. However, the mechanism behind the improvement from SRA to Self-Flow, dual-time scheduling, remains under-examined: Self-Flow attributes its gain to interactions between tokens at different noise levels, where cleaner tokens help infer noisier ones. In this work, we revisit this explanation and ask whether the gain instead comes from data augmentation along the noise dimension. To disentangle these factors, we introduce Attention Separation, which preserves the same dual-timestep input as Self-Flow while blocking attention between tokens assigned to different noise levels. Surprisingly, removing such interaction does not degrade performance and can even improve it, suggesting that the improvement from SRA to Self-Flow mainly comes from data augmentation. Furthermore,We show that Attention Separation itself provides an augmentation effect by splitting a single image into multiple effective training parts to expand the training data. Based on these observations, we combine self-representation alignment with dual-timestep and attention-separation augmentation, and demonstrate the effectiveness of this design on ImageNet.

📄 PDF Abstract BibTeX arXiv:2607.02508

Code (0)

등록된 구현이 없습니다.

Tasks

Data Augmentation

Similar Papers 제목 키워드 기반

Learning by Analogy: Reliable Supervision from Transformations for Unsupervised Optical Flow Estimation

2020-03-29 · CVPR 2020 6 · Liang Liu, Jiangning Zhang, Ruifei He, Yong liu 외

Unsupervised learning of optical flow, which leverages the supervision from view synthesis, has emerged as a promising alternative to supervised methods. However, the objective of unsupervised learning is likely to be un…

DecoderOptical Flow EstimationSelf-Supervised Learning

Cross Pixel Optical Flow Similarity for Self-Supervised Learning

2018-07-15 · Aravindh Mahendran, James Thewlis, Andrea Vedaldi

We propose a novel method for learning convolutional neural image representations without manual supervision. We use motion cues in the form of optical flow, to supervise representations of static images. The obvious app…

image-classificationImage ClassificationImage SegmentationOptical Flow Estimation+3

Self-Supervision in Time for Satellite Images(S3-TSS): A novel method of SSL technique in Satellite images

2024-03-07 · Akansh Maurya, Hewan Shrestha, Mohammad Munem Shahriar

With the limited availability of labeled data with various atmospheric conditions in remote sensing images, it seems useful to work with self-supervised algorithms. Few pretext-based algorithms, including from rotation, …

Self-Supervised Learning

Equivariant Spatio-Temporal Self-Supervision for LiDAR Object Detection

2024-04-17 · Deepti Hegde, Suhas Lohit, Kuan-Chuan Peng, Michael J. Jones 외

Popular representation learning methods encourage feature invariance under transformations applied at the input. However, in 3D perception tasks like object localization and segmentation, outputs are naturally equivarian…

3D Object DetectionObjectobject-detectionObject Detection+2

FAFA: Frequency-Aware Flow-Aided Self-Supervision for Underwater Object Pose Estimation

2024-09-25 · Jingyi Tang, Gu Wang, Zeyu Chen, Shengquan Li 외

Although methods for estimating the pose of objects in indoor scenes have achieved great success, the pose estimation of underwater objects remains challenging due to difficulties brought by the complex underwater enviro…

6D Pose EstimationPose Estimation