A Comprehensive Study of Knowledge Editing for Large Language Models

Ningyu Zhang, Yunzhi Yao, Bozhong Tian, Peng Wang, Shumin Deng, Mengru Wang, Zekun Xi, Shengyu Mao, Jintian Zhang, Yuansheng Ni, Siyuan Cheng, Ziwen Xu, Xin Xu, Jia-Chen Gu, Yong Jiang, Pengjun Xie, Fei Huang, Lei Liang, Zhiqiang Zhang, Xiaowei Zhu, Jun Zhou, Huajun Chen

2024-01-02Model Editing knowledge editing

Paper PDF Code(official)Code(official)

Abstract

Large Language Models (LLMs) have shown extraordinary capabilities in understanding and generating text that closely mirrors human communication. However, a primary limitation lies in the significant computational demands during training, arising from their extensive parameterization. This challenge is further intensified by the dynamic nature of the world, necessitating frequent updates to LLMs to correct outdated information or integrate new knowledge, thereby ensuring their continued relevance. Note that many applications demand continual model adjustments post-training to address deficiencies or undesirable behaviors. There is an increasing interest in efficient, lightweight methods for on-the-fly model modifications. To this end, recent years have seen a burgeoning in the techniques of knowledge editing for LLMs, which aim to efficiently modify LLMs' behaviors within specific domains while preserving overall performance across various inputs. In this paper, we first define the knowledge editing problem and then provide a comprehensive review of cutting-edge approaches. Drawing inspiration from educational and cognitive research theories, we propose a unified categorization criterion that classifies knowledge editing methods into three groups: resorting to external knowledge, merging knowledge into the model, and editing intrinsic knowledge. Furthermore, we introduce a new benchmark, KnowEdit, for a comprehensive empirical evaluation of representative knowledge editing approaches. Additionally, we provide an in-depth analysis of knowledge location, which can give a deeper understanding of the knowledge structures inherent within LLMs. Finally, we discuss several potential applications of knowledge editing, outlining its broad and impactful implications.

Results

Task	Dataset	Metric	Value	Model
Video	zsRE	edit success	96.74	MEND
Video	zsRE	fluency	586.34	MEND
Video	zsRE	locality	92.87	MEND
Video	zsRE	portability	60.41	MEND
Video	zsRE	edit success	96.57	ROME
Video	zsRE	locality	27.14	ROME
Video	zsRE	portability	52.2	ROME
Temporal Action Localization	zsRE	edit success	96.74	MEND
Temporal Action Localization	zsRE	fluency	586.34	MEND
Temporal Action Localization	zsRE	locality	92.87	MEND
Temporal Action Localization	zsRE	portability	60.41	MEND
Temporal Action Localization	zsRE	edit success	96.57	ROME
Temporal Action Localization	zsRE	locality	27.14	ROME
Temporal Action Localization	zsRE	portability	52.2	ROME
Zero-Shot Learning	zsRE	edit success	96.74	MEND
Zero-Shot Learning	zsRE	fluency	586.34	MEND
Zero-Shot Learning	zsRE	locality	92.87	MEND
Zero-Shot Learning	zsRE	portability	60.41	MEND
Zero-Shot Learning	zsRE	edit success	96.57	ROME
Zero-Shot Learning	zsRE	locality	27.14	ROME
Zero-Shot Learning	zsRE	portability	52.2	ROME
Activity Recognition	zsRE	edit success	96.74	MEND
Activity Recognition	zsRE	fluency	586.34	MEND
Activity Recognition	zsRE	locality	92.87	MEND
Activity Recognition	zsRE	portability	60.41	MEND
Activity Recognition	zsRE	edit success	96.57	ROME
Activity Recognition	zsRE	locality	27.14	ROME
Activity Recognition	zsRE	portability	52.2	ROME
Action Localization	zsRE	edit success	96.74	MEND
Action Localization	zsRE	fluency	586.34	MEND
Action Localization	zsRE	locality	92.87	MEND
Action Localization	zsRE	portability	60.41	MEND
Action Localization	zsRE	edit success	96.57	ROME
Action Localization	zsRE	locality	27.14	ROME
Action Localization	zsRE	portability	52.2	ROME
3D Action Recognition	zsRE	edit success	96.74	MEND
3D Action Recognition	zsRE	fluency	586.34	MEND
3D Action Recognition	zsRE	locality	92.87	MEND
3D Action Recognition	zsRE	portability	60.41	MEND
3D Action Recognition	zsRE	edit success	96.57	ROME
3D Action Recognition	zsRE	locality	27.14	ROME
3D Action Recognition	zsRE	portability	52.2	ROME
Action Recognition	zsRE	edit success	96.74	MEND
Action Recognition	zsRE	fluency	586.34	MEND
Action Recognition	zsRE	locality	92.87	MEND
Action Recognition	zsRE	portability	60.41	MEND
Action Recognition	zsRE	edit success	96.57	ROME
Action Recognition	zsRE	locality	27.14	ROME
Action Recognition	zsRE	portability	52.2	ROME
Model Editing	zsRE	edit success	96.74	MEND
Model Editing	zsRE	fluency	586.34	MEND
Model Editing	zsRE	locality	92.87	MEND
Model Editing	zsRE	portability	60.41	MEND
Model Editing	zsRE	edit success	96.57	ROME
Model Editing	zsRE	locality	27.14	ROME
Model Editing	zsRE	portability	52.2	ROME

A Comprehensive Study of Knowledge Editing for Large Language Models

Abstract

Results

Task	Dataset	Metric	Value	Model
Video	zsRE	edit success	96.74	MEND
Video	zsRE	fluency	586.34	MEND
Video	zsRE	locality	92.87	MEND
Video	zsRE	portability	60.41	MEND
Video	zsRE	edit success	96.57	ROME
Video	zsRE	locality	27.14	ROME
Video	zsRE	portability	52.2	ROME
Temporal Action Localization	zsRE	edit success	96.74	MEND
Temporal Action Localization	zsRE	fluency	586.34	MEND
Temporal Action Localization	zsRE	locality	92.87	MEND
Temporal Action Localization	zsRE	portability	60.41	MEND
Temporal Action Localization	zsRE	edit success	96.57	ROME
Temporal Action Localization	zsRE	locality	27.14	ROME
Temporal Action Localization	zsRE	portability	52.2	ROME
Zero-Shot Learning	zsRE	edit success	96.74	MEND
Zero-Shot Learning	zsRE	fluency	586.34	MEND
Zero-Shot Learning	zsRE	locality	92.87	MEND
Zero-Shot Learning	zsRE	portability	60.41	MEND
Zero-Shot Learning	zsRE	edit success	96.57	ROME
Zero-Shot Learning	zsRE	locality	27.14	ROME
Zero-Shot Learning	zsRE	portability	52.2	ROME
Activity Recognition	zsRE	edit success	96.74	MEND
Activity Recognition	zsRE	fluency	586.34	MEND
Activity Recognition	zsRE	locality	92.87	MEND
Activity Recognition	zsRE	portability	60.41	MEND
Activity Recognition	zsRE	edit success	96.57	ROME
Activity Recognition	zsRE	locality	27.14	ROME
Activity Recognition	zsRE	portability	52.2	ROME
Action Localization	zsRE	edit success	96.74	MEND
Action Localization	zsRE	fluency	586.34	MEND
Action Localization	zsRE	locality	92.87	MEND
Action Localization	zsRE	portability	60.41	MEND
Action Localization	zsRE	edit success	96.57	ROME
Action Localization	zsRE	locality	27.14	ROME
Action Localization	zsRE	portability	52.2	ROME
3D Action Recognition	zsRE	edit success	96.74	MEND
3D Action Recognition	zsRE	fluency	586.34	MEND
3D Action Recognition	zsRE	locality	92.87	MEND
3D Action Recognition	zsRE	portability	60.41	MEND
3D Action Recognition	zsRE	edit success	96.57	ROME
3D Action Recognition	zsRE	locality	27.14	ROME
3D Action Recognition	zsRE	portability	52.2	ROME
Action Recognition	zsRE	edit success	96.74	MEND
Action Recognition	zsRE	fluency	586.34	MEND
Action Recognition	zsRE	locality	92.87	MEND
Action Recognition	zsRE	portability	60.41	MEND
Action Recognition	zsRE	edit success	96.57	ROME
Action Recognition	zsRE	locality	27.14	ROME
Action Recognition	zsRE	portability	52.2	ROME
Model Editing	zsRE	edit success	96.74	MEND
Model Editing	zsRE	fluency	586.34	MEND
Model Editing	zsRE	locality	92.87	MEND
Model Editing	zsRE	portability	60.41	MEND
Model Editing	zsRE	edit success	96.57	ROME
Model Editing	zsRE	locality	27.14	ROME
Model Editing	zsRE	portability	52.2	ROME

A Comprehensive Study of Knowledge Editing for Large Language Models

Abstract

Results

Related Papers

A Comprehensive Study of Knowledge Editing for Large Language Models

Abstract

Results

Related Papers