paper-with-me

홈 › Papers

Training Large Language Models for Advanced Typosquatting Detection

2025-03-28 · Jackson Welch

Typosquatting is a long-standing cyber threat that exploits human error in typing URLs to deceive users, distribute malware, and conduct phishing attacks. With the proliferation of domain names and new Top-Level Domains (TLDs), typosquatting techniques have grown more sophisticated, posing significant risks to individuals, businesses, and national cybersecurity infrastructure. Traditional detection methods primarily focus on well-known impersonation patterns, leaving gaps in identifying more complex attacks. This study introduces a novel approach leveraging large language models (LLMs) to enhance typosquatting detection. By training an LLM on character-level transformations and pattern-based heuristics rather than domain-specific data, a more adaptable and resilient detection mechanism develops. Experimental results indicate that the Phi-4 14B model outperformed other tested models when properly fine tuned achieving a 98% accuracy rate with only a few thousand training samples. This research highlights the potential of LLMs in cybersecurity applications, specifically in mitigating domain-based deception tactics, and provides insights into optimizing machine learning strategies for threat detection.

📄 PDF Abstract BibTeX arXiv:2503.22406

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

DNS Typo-squatting Domain Detection: A Data Analytics & Machine Learning Based Approach

2020-12-25 · Abdallah Moubayed, MohammadNoor Injadat, Abdallah Shami, Hanan Lutfiyya

Domain Name System (DNS) is a crucial component of current IP-based networks as it is the standard mechanism for name to IP resolution. However, due to its lack of data integrity and origin authentication processes, it i…

BIG-bench Machine LearningClusteringEnsemble Learning

Enhancing Sentiment Classification and Irony Detection in Large Language Models through Advanced Prompt Engineering Techniques

2026-01-13 · Marvin Schmitt, Anne Schwerk, Sebastian Lempert arxiv

This study investigates the use of prompt engineering to enhance large language models (LLMs), specifically GPT-4o-mini and gemini-1.5-flash, in sentiment analysis tasks. It evaluates advanced prompting techniques like f…

Sentiment AnalysisPrompt EngineeringFew-Shot Learning

Hidden Entity Detection from GitHub Leveraging Large Language Models

2025-01-08 · Lu Gan, Martin Blum, Danilo Dessi, Brigitte Mathiak 외

Named entity recognition is an important task when constructing knowledge bases from unstructured data sources. Whereas entity detection methods mostly rely on extensive training data, Large Language Models (LLMs) have p…

Few-Shot Learningnamed-entity-recognitionNamed Entity RecognitionPrompt Learning+1

LLMcap: Large Language Model for Unsupervised PCAP Failure Detection

2024-07-03 · Lukasz Tulczyjew, Kinan Jarrah, Charles Abondo, Dina Bennett 외

The integration of advanced technologies into telecommunication networks complicates troubleshooting, posing challenges for manual error identification in Packet Capture (PCAP) data. This manual approach, requiring subst…

Language ModelingLanguage ModellingLarge Language ModelMasked Language Modeling+1

Advanced Deep Learning and Large Language Models: Comprehensive Insights for Cancer Detection

2025-03-30 · Yassine Habchi, Hamza Kheddar, Yassine Himeur, Adel Belouchrani 외

The rapid advancement of deep learning (DL) has transformed healthcare, particularly in cancer detection and diagnosis. DL surpasses traditional machine learning and human accuracy, making it a critical tool for identify…

DiagnosticFederated LearningReinforcement Learning (RL)Transfer Learning