paper-with-me

홈 › Papers

DRIFT: Drift-Resilient Invariant-Feature Transformer for DGA Detection

2026-05-11 · Chaeyoung Lee, Chaeri Jung, Seonghoon Jeong arxiv

Domain Generation Algorithms (DGAs) evolve continuously to evade botnet detection, posing a persistent challenge for dependable network defense. While deep learning-based detectors achieve strong performance under static conditions, they suffer severe degradation when facing temporal drift. Through a 9-year longitudinal study (2017-2025), we empirically show that state-of-the-art character- and word-based DGA classifiers rapidly lose effectiveness as new DGA variants emerge. To address this problem, we propose a drift-resilient Transformer-based framework that learns invariant representations through a hybrid tokenization strategy and multi-task self-supervised pre-training. The model integrates (i) character-level encoding to capture stochastic morphological patterns and (ii) subword-level encoding for word-based DGAs. Three pre-training tasks enable the model to learn robust structural and contextual features prior to supervised fine-tuning. Comprehensive evaluations demonstrate that our method significantly mitigates temporal degradation and consistently outperforms state-of-the-art baselines in forward-chaining experiments. The proposed approach offers a dependable foundation for long-term DGA defense in evolving threat landscapes. Our code is available at: https://github.com/snsec-net/2026-DSN-DRIFT.

📄 PDF Abstract BibTeX arXiv:2605.10436

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Explaining Drift using Shapley Values

2024-01-18 · Narayanan U. Edakunni, Utkarsh Tekriwal, Anukriti Jain

Machine learning models often deteriorate in their performance when they are used to predict the outcomes over data on which they were not trained. These scenarios can often arise in real world when the distribution of d…

Rolling with the Punches: Resilient Contrastive Pre-training under Non-Stationary Drift

2025-02-11 · Xiaoyu Yang, Jie Lu, En Yu

The remarkable success of large-scale contrastive pre-training, fueled by vast and curated datasets, is encountering new frontiers as the scaling paradigm evolves. A critical emerging challenge is the effective pre-train…

Causal InferenceContrastive Learning

Optimized Deep Learning Models for Malware Detection under Concept Drift

2023-08-21 · William Maillet, Benjamin Marais

Despite the promising results of machine learning models in malicious files detection, they face the problem of concept drift due to their constant evolution. This leads to declining performance over time, as the data di…

Malware Detection

Drift-Aware Federated Learning: A Causal Perspective

2025-03-12 · Yunjie Fang, Sheng Wu, Tao Yang, Xiaofeng Wu 외

Federated learning (FL) facilitates collaborative model training among multiple clients while preserving data privacy, often resulting in enhanced performance compared to models trained by individual clients. However, fa…

Federated Learning

SymDrift: One-Shot Generative Modeling under Symmetries

2026-05-07 · Samir Darouich, Vinh Tong, Lluís Pastor-Pérez, Tanja Bien 외 arxiv

Generative modeling of physical systems, such as molecules, requires learning distributions that are invariant under global symmetries, such as rotations in three-dimensional space. Equivariant diffusion and flow matchin…