paper-with-me

Papers Model Editing

“Model Editing” 태그가 달린 논문 193편 · 필터 해제

Mitigating Gender Bias in Code Large Language Models via Model Editing

2024-10-10 · Zhanyue Qin, Haochuan Wang, Zecheng Wang, Deyuan Liu 외

In recent years, with the maturation of large language model (LLM) technology and the emergence of high-quality programming code datasets, researchers have become increasingly confident in addressing the challenges of pr…

Code Generationknowledge editingLarge Language ModelModel Editing+1

Uncovering Overfitting in Large Language Model Editing

2024-10-10 · Mengqi Zhang, Xiaotian Ye, Qiang Liu, Pengjie Ren 외

Knowledge editing has been proposed as an effective method for updating and correcting the internal knowledge of Large Language Models (LLMs). However, existing editing methods often struggle with complex tasks, such as …

AttributeIn-Context Learningknowledge editingLanguage Modeling+4

Mitigating the Language Mismatch and Repetition Issues in LLM-based Machine Translation via Model Editing

2024-10-09 · Weichuan Wang, Zhaoyi Li, Defu Lian, Chen Ma 외

Large Language Models (LLMs) have recently revolutionized the NLP field, while they still fall short in some specific down-stream tasks. In the work, we focus on utilizing LLMs to perform machine translation, where we ob…

Machine TranslationModel EditingTranslation

FAME: Towards Factual Multi-Task Model Editing

2024-10-07 · Li Zeng, Yingyu Shan, Zeming Liu, Jiashu Yao 외

Large language models (LLMs) embed extensive knowledge and utilize it to perform exceptionally well across various tasks. Nevertheless, outdated knowledge or factual errors within LLMs can lead to misleading or incorrect…

modelModel Editing

Neuron-Level Sequential Editing for Large Language Models

2024-10-05 · Houcheng Jiang, Junfeng Fang, Tianyu Zhang, An Zhang 외

This work explores sequential model editing in large language models (LLMs), a critical task that involves modifying internal knowledge within LLMs continuously through multi-round editing, each incorporating updates or …

Model Editing

AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models

2024-10-03 · Junfeng Fang, Houcheng Jiang, Kun Wang, Yunshan Ma 외

Large language models (LLMs) often exhibit hallucinations due to incorrect or outdated knowledge. Hence, model editing methods have emerged to enable targeted knowledge updates. To achieve this, a prevailing paradigm is …

knowledge editingModel Editing

Better Call SAUL: Fluent and Consistent Language Model Editing with Generation Regularization

2024-10-03 · Mingyang Wang, Lukas Lange, Heike Adel, Jannik Strötgen 외

To ensure large language models contain up-to-date knowledge, they need to be updated regularly. However, model editing is challenging as it might also affect knowledge that is unrelated to the new data. State-of-the-art…

Language ModelingLanguage ModellingModel EditingSentence

UniAdapt: A Universal Adapter for Knowledge Calibration

2024-10-01 · Tai D. Nguyen, Long H. Pham, Jun Sun

Large Language Models (LLMs) require frequent updates to correct errors and keep pace with continuously evolving knowledge in a timely and effective manner. Recent research in it model editing has highlighted the challen…

Mixture-of-ExpertsModel EditingRetrieval-augmented GenerationSemantic Similarity+1

Self-Updatable Large Language Models with Parameter Integration

2024-10-01 · Yu Wang, Xinshuang Liu, Xiusi Chen, Sean O'Brien 외

Despite significant advancements in large language models (LLMs), the rapid and frequent integration of small-scale experiences, such as interactions with surrounding objects, remains a substantial challenge. Two critica…

Continual LearningConversational RecommendationKnowledge DistillationModel Editing+1

"Why" Has the Least Side Effect on Model Editing

2024-09-27 · Tsung-Hsuan Pan, Chung-Chi Chen, Hen-Hsen Huang, Hsin-Hsi Chen

Training large language models (LLMs) from scratch is an expensive endeavor, particularly as world knowledge continually evolves. To maintain relevance and accuracy of LLMs, model editing has emerged as a pivotal researc…

Experimental Designknowledge editingModel EditingWorld Knowledge

Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis

2024-09-21 · Zeping Yu, Sophia Ananiadou

We find arithmetic ability resides within a limited number of attention heads, with each head specializing in distinct operations. To delve into the reason, we introduce the Comparative Neuron Analysis (CNA) method, whic…

Model EditingPrediction

Property Neurons in Self-Supervised Speech Transformers

2024-09-07 · Tzu-Quan Lin, Guan-Ting Lin, Hung-Yi Lee, Hao Tang

There have been many studies on analyzing self-supervised speech Transformers, in particular, with layer-wise analysis. It is, however, desirable to have an approach that can pinpoint exactly a subset of neurons that is …

Model Editing

TimeDiT: General-purpose Diffusion Transformers for Time Series Foundation Model

2024-09-03 · Defu Cao, Wen Ye, Yizhou Zhang, Yan Liu

With recent advances in building foundation models for texts and video data, there is a surge of interest in foundation models for time series. A family of models have been developed, utilizing a temporal auto-regressive…

Anomaly DetectionDenoisingImputationMissing Values+2

Enhancing Multi-hop Reasoning through Knowledge Erasure in Large Language Model Editing

2024-08-22 · Mengqi Zhang, Bowen Fang, Qiang Liu, Pengjie Ren 외

Large language models (LLMs) face challenges with internal knowledge inaccuracies and outdated information. Knowledge editing has emerged as a pivotal approach to mitigate these issues. Although current knowledge editing…

knowledge editingLanguage ModelingLanguage ModellingLarge Language Model+1

MEGen: Generative Backdoor in Large Language Models via Model Editing

2024-08-20 · Jiyang Qiu, Xinbei Ma, Zhuosheng Zhang, Hai Zhao

Large language models (LLMs) have demonstrated remarkable capabilities. Their powerful generative abilities enable flexible responses based on various queries or instructions. Emerging as widely adopted generalists for d…

Backdoor AttackLanguage ModellingModel Editing

Promoting Equality in Large Language Models: Identifying and Mitigating the Implicit Bias based on Bayesian Theory

2024-08-20 · Yongxin Deng, Xihe Qiu, Xiaoyu Tan, Jing Pan 외

Large language models (LLMs) are trained on extensive text corpora, which inevitably include biased information. Although techniques such as Affective Alignment can mitigate some negative impacts of these biases, existin…

Model Editing

Attribution Analysis Meets Model Editing: Advancing Knowledge Correction in Vision Language Models with VisEdit

2024-08-19 · Qizhou Chen, Taolin Zhang, Chengyu Wang, Xiaofeng He 외

Model editing aims to correct outdated or erroneous knowledge in large models without costly retraining. Recent research discovered that the mid-layer representation of the subject's final token in a prompt has a strong …

DecoderLanguage ModelingLanguage ModellingLarge Language Model+1

Resolving Lexical Bias in Edit Scoping with Projector Editor Networks

2024-08-19 · Hammad Rizwan, Domenic Rosati, Ga Wu, Hassan Sajjad

Weight-preserving model editing techniques heavily rely on the scoping mechanism that decides when to apply an edit to the base model. These scoping mechanisms utilize distance functions in the representation space to as…

Contrastive LearningModel Editing

ELDER: Enhancing Lifelong Model Editing with Mixture-of-LoRA

2024-08-19 · Jiaang Li, Quan Wang, Zhongnan Wang, Yongdong Zhang 외

Large language models (LLMs) require model editing to efficiently update specific knowledge within them and avoid factual errors. Most model editing methods are solely designed for single-time use and result in a signifi…

Model Editing

Can LLMs be Fooled? Investigating Vulnerabilities in LLMs

2024-07-30 · Sara Abdali, Jia He, CJ Barberan, Richard Anarfi

The advent of Large Language Models (LLMs) has garnered significant popularity and wielded immense power across various domains within Natural Language Processing (NLP). While their capabilities are undeniably impressive…

Model Editing
← 이전 61–80 / 193 다음 →