paper-with-me

Papers Memorization

“Memorization” 태그가 달린 논문 1,088편 · 필터 해제

What Should LLMs Forget? Quantifying Personal Data in LLMs for Right-to-Be-Forgotten Requests

2025-07-15 · Dimitri Staufer

Large Language Models (LLMs) can memorize and reveal personal information, raising concerns regarding compliance with the EU's GDPR, particularly the Right to Be Forgotten (RTBF). Existing machine unlearning methods assu…

Machine UnlearningMemorization

Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination

2025-07-14 · Mingqi Wu, Zhihao Zhang, Qiaole Dong, Zhiheng Xi 외

The reasoning capabilities of large language models (LLMs) have been a longstanding focus of research. Recent works have further enhanced these capabilities using reinforcement learning (RL), with many new methods claimi…

MathMathematical ReasoningMemorizationReinforcement Learning (RL)

Entropy-Memorization Law: Evaluating Memorization Difficulty of Data in LLMs

2025-07-08 · Yizhan Huang, Zhe Yang, Meifang Chen, Jianping Zhang 외

Large Language Models (LLMs) are known to memorize portions of their training data, sometimes reproducing content verbatim when prompted appropriately. In this work, we investigate a fundamental yet under-explored questi…

Memorization

MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI

2025-06-30 · Huanjin Yao, Jiaxing Huang, Yawen Qiu, Michael K. Chen 외

Reasoning plays a crucial role in advancing Multimodal Large Language Models (MLLMs) toward Artificial General Intelligence. However, existing MLLM benchmarks often fall short in precisely and comprehensively evaluating …

Memorization

Listener-Rewarded Thinking in VLMs for Image Preferences

2025-06-28 · Alexander Gambashidze, Li Pengyi, Matvey Skripkin, Andrey Galichin 외

Training robust and generalizable reward models for human visual preferences is essential for aligning text-to-image and text-to-video generative models with human intent. However, current reward models often fail to gen…

MemorizationReinforcement Learning (RL)

Where to find Grokking in LLM Pretraining? Monitor Memorization-to-Generalization without Test

2025-06-26 · Ziyue Li, Chenrui Fan, Tianyi Zhou

Grokking, i.e., test performance keeps improving long after training loss converged, has been recently witnessed in neural network training, making the mechanism of generalization and other emerging capabilities such as …

Code GenerationLarge Language ModelMathMemorization

Counterfactual Influence as a Distributional Quantity

2025-06-25 · Matthieu Meeus, Igor Shilov, Georgios Kaissis, Yves-Alexandre de Montjoye

Machine learning models are known to memorize samples from their training data, raising concerns around privacy and generalization. Counterfactual self-influence is a popular metric to study memorization, quantifying how…

counterfactualimage-classificationImage ClassificationMemorization+1

Leaner Training, Lower Leakage: Revisiting Memorization in LLM Fine-Tuning with LoRA

2025-06-25 · Fei Wang, Baochun Li

Memorization in large language models (LLMs) makes them vulnerable to data extraction attacks. While pre-training memorization has been extensively studied, fewer works have explored its impact in fine-tuning, particular…

Memorization

Uncovering Conceptual Blindspots in Generative Image Models Using Sparse Autoencoders

2025-06-24 · Matyas Bohacek, Thomas Fel, Maneesh Agrawala, Ekdeep Singh Lubana

Despite their impressive performance, generative image models trained on large-scale datasets frequently fail to produce images with seemingly simple concepts -- e.g., human hands or objects appearing in groups of four -…

Memorization

Robots and Children that Learn Together : Improving Knowledge Retention by Teaching Peer-Like Interactive Robots

2025-06-23 · Imene Tarakli, Samuele Vinanzi, Richard Moore, Alessandro Di Nuovo

Despite growing interest in Learning-by-Teaching (LbT), few studies have explored how this paradigm can be implemented with autonomous, peer-like social robots in real classrooms. Most prior work has relied on scripted o…

MemorizationReinforcement Learning (RL)

A Random Matrix Analysis of In-context Memorization for Nonlinear Attention

2025-06-23 · Zhenyu Liao, Jiaqing Liu, Tianqi Hou, Difan Zou 외

Attention mechanisms have revolutionized machine learning (ML) by enabling efficient modeling of global dependencies across inputs. Their inherently parallelizable structures allow for efficient scaling with the exponent…

Memorization

In-Context Learning Strategies Emerge Rationally

2025-06-21 · Daniel Wurgaft, Ekdeep Singh Lubana, Core Francisco Park, Hidenori Tanaka 외

Recent work analyzing in-context learning (ICL) has identified a broad set of strategies that describe model behavior in different experimental conditions. We aim to unify these findings by asking why a model learns thes…

In-Context LearningMemorization

LexiMark: Robust Watermarking via Lexical Substitutions to Enhance Membership Verification of an LLM's Textual Training Data

2025-06-17 · Eyal German, Sagiv Antebi, Edan Habler, Asaf Shabtai 외

Large language models (LLMs) can be trained or fine-tuned on data obtained without the owner's consent. Verifying whether a specific LLM was trained on particular data instances or an entire dataset is extremely challeng…

Memorization

Capacity Matters: a Proof-of-Concept for Transformer Memorization on Real-World Data

2025-06-17 · Anton Changalidis, Aki Härmä

This paper studies how the model architecture and data configurations influence the empirical memorization capacity of generative transformers. The models are trained using synthetic text datasets derived from the System…

Memorization

Less is More: Undertraining Experts Improves Model Upcycling

2025-06-17 · Stefan Horoi, Guy Wolf, Eugene Belilovsky, Gintare Karolina Dziugaite

Modern deep learning is increasingly characterized by the use of open-weight foundation models that can be fine-tuned on specialized datasets. This has led to a proliferation of expert models and adapters, often shared v…

MemorizationmodelTransfer Learning

Dataset distillation for memorized data: Soft labels can leak held-out teacher knowledge

2025-06-17 · Freya Behrens, Lenka Zdeborová

Dataset distillation aims to compress training data into fewer examples via a teacher, from which a student can learn effectively. While its success is often attributed to structure in the data, modern neural networks al…

Dataset DistillationMemorization

Winter Soldier: Backdooring Language Models at Pre-Training with Indirect Data Poisoning

2025-06-17 · Wassim Bouaziz, Mathurin Videau, Nicolas Usunier, El-Mahdi El-Mhamdi

The pre-training of large language models (LLMs) relies on massive text datasets sourced from diverse and difficult-to-curate origins. Although membership inference attacks and hidden canaries have been explored to trace…

Data PoisoningMemorization

Sharpness-Aware Machine Unlearning

2025-06-16 · Haoran Tang, Rajiv Khanna

We characterize the effectiveness of Sharpness-aware minimization (SAM) under machine unlearning scheme, where unlearning forget signals interferes with learning retain signals. While previous work prove that SAM improve…

DenoisingMachine UnlearningMemorization

Restoring Gaussian Blurred Face Images for Deanonymization Attacks

2025-06-14 · Haoyu Zhai, Shuo Wang, Pirouz Naghavi, Qingying Hao 외

Gaussian blur is widely used to blur human faces in sensitive photos before the photos are posted on the Internet. However, it is unclear to what extent the blurred faces can be restored and used to re-identify the perso…

DeblurringFace AnonymizationMemorization

The SWE-Bench Illusion: When State-of-the-Art LLMs Remember Instead of Reason

2025-06-14 · Shanchao Liang, Spandan Garg, Roshanak Zilouchian Moghaddam

As large language models (LLMs) become increasingly capable and widely adopted, benchmarks play a central role in assessing their practical utility. For example, SWE-Bench Verified has emerged as a critical benchmark for…

DiagnosticMemorization
1–20 / 1,088 다음 →