Published December 2024 | Version v1
Dissertation Open

Continual Learning and Knowledge Editing of Large Language Models with Relevance-Based Parameter Activation

  • 1. University of Chicago

Contributors

Advisor:

Description

Knowledge editing in Large Language Models (LLMs) aims to make precise updates to specific pieces of information, correcting inaccuracies or biases without unintentionally altering unrelated knowledge or skills. This field of research addresses three essential challenges: generalization (how well the model applies edited information across various contexts), locality (making accurate changes without impacting unrelated information), and scalability (ensuring performance remains efficient as the number of edits increases). Our research introduces two key contributions: (1) a new knowledge editing benchmark that overcomes limitations in existing benchmarks, providing materials suitable for fine-tuning and thorough evaluations, and (2) a novel approach using external memory to manage knowledge edits. This approach, called Relevance-based Parameter Activation (rel-par-act), utilizes an embedding model and vector store to activate LoRA layers tailored to specific edits. Our method achieves state-of-the-art performance in both generalization and locality on our benchmark and can scale to hundreds of edits with high efficiency.

Files

knowledge_editing_rpa.pdf

Files (2.1 MB)

Name Size Download all
Dissertation
md5:a9b467d810f0122da2c35a72c6f84b37
2.1 MB Preview Download

Additional details

Identifiers

Other
oai:uchicago.tind.io:13996

UChicago Information

Division(s)
Physical Sciences Division
Department(s)
Computer Science