paper-with-me

Papers

Assessing Deanonymization Risks with Stylometry-Assisted LLM Agent

2026-02-26 · Boyang Zhang, Yang Zhang arxiv

The rapid advancement of large language models (LLMs) has enabled powerful authorship inference capabilities, raising growing concerns about unintended deanonymization risks in textual data such as news articles. In this work, we introduce an LLM agent designed to evaluate and mitigate such risks through a structured, interpretable pipeline. Central to our framework is the proposed $\textit{SALA}$ (Stylometry-Assisted LLM Analysis) method, which integrates quantitative stylometric features with LLM reasoning for robust and transparent authorship attribution. Experiments on large-scale news datasets demonstrate that $\textit{SALA}$, particularly when augmented with a database module, achieves high inference accuracy in various scenarios. Finally, we propose a guided recomposition strategy that leverages the agent's reasoning trace to generate rewriting prompts, effectively reducing authorship identifiability while preserving textual meaning. Our findings highlight both the deanonymization potential of LLM agents and the importance of interpretable, proactive defenses for safeguarding author privacy.

📄 PDF Abstract BibTeX arXiv:2602.23079

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Reproduction and Replication of an Adversarial Stylometry Experiment

2022-08-15 · Haining Wang, Patrick Juola, Allen Riddell

Maintaining anonymity while communicating using natural language remains a challenge. Standard authorship attribution techniques that analyze candidate authors' writing styles achieve uncomfortably high accuracy even whe…

Authorship AttributionTranslation

Gradient-Leaks: Understanding and Controlling Deanonymization in Federated Learning

2018-05-15 · Tribhuvanesh Orekondy, Seong Joon Oh, Yang Zhang, Bernt Schiele 외

Federated Learning (FL) systems are gaining popularity as a solution to training Machine Learning (ML) models from large-scale user data collected on personal devices (e.g., smartphones) without their raw data leaving th…

Data AugmentationFederated LearningSpeech Recognition

Password-conditioned Anonymization and Deanonymization with Face Identity Transformers

2020-08-01 · ECCV 2020 8 · Xiuye Gu, Weixin Luo, Michael S. Ryoo, Yong Jae Lee

Cameras are prevalent in our daily lives, and enable many useful systems built upon computer vision technologies such as smart cameras and home robots for service applications. However, there is also an increasing societ…

Multi-Task Learning

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents

2025-07-01 · Siyuan Liang, Tianmeng Fang, Zhe Liu, Aishan Liu 외 arxiv

With the wide application of multimodal foundation models in intelligent agent systems, scenarios such as mobile device control, intelligent assistant interaction, and multimodal task execution are gradually relying on s…

On the Simultaneous Preservation of Privacy and Community Structure in Anonymized Networks

2016-03-25 · Daniel Cullina, Kushagra Singhal, Negar Kiyavash, Prateek Mittal

We consider the problem of performing community detection on a network, while maintaining privacy, assuming that the adversary has access to an auxiliary correlated network. We ask the question "Does there exist a regime…

Community DetectionStochastic Block Model