Editing Factual Knowledge in Language Models

The factual knowledge acquired during pretraining and stored in the parameters of Language Models (LMs) can be useful in downstream tasks (e.g., question answering or textual inference). However, some facts can be incorrectly induced or become obsolete over time. We present KNOWLEDGEEDITOR, a method which can be used to edit this knowledge and, thus, fix 'bugs' or unexpected predictions without the need for expensive retraining or fine-tuning. Besides being computationally efficient, KNOWLEDGEEDITOR does not require any modifications in LM pretraining (e.g., the use of meta-learning). In our approach, we train a hyper-network with constrained optimization to modify a fact without affecting the rest of the knowledge; the trained hyper-network is then used to predict the weight update at test time. We show KNOWL-EDGEEDITOR's efficacy with two popular architectures and knowledge-intensive tasks: i) a BERT model fine-tuned for fact-checking, and ii) a sequence-to-sequence BART model for question answering. With our method, changing a prediction on the specific wording of a query tends to result in a consistent change in predictions also for its paraphrases. We show that this can be further encouraged by exploiting (e.g., automatically-generated) paraphrases during training. Interestingly, our hyper-network can be regarded as a 'probe' revealing which components need to be changed to manipulate factual knowledge; our analysis shows that the updates tend to be concentrated on a small subset of components. 1 How is Namibia's capital city called? Semantically equivalent Answers Scores Namibia Nigeria Nibia Namibia Tasman -0.43 -0.69 -0.89 -1.08 -1.19 What is the capital of Namibia? Answers Scores Namibia Nigeria Nibia Tasman Namibia -0.32 -0.79 -0.87 -1.14 -1.16 What is the capital of Russia? Answers Scores

Long Short-Term MemoryLong Short-Term MemoryConvex OptimizationConvex OptimizationSequence to SequenceLearning with Neural…Sequence to Sequence Learning with Neural NetworksDropout: a simple way toprevent neural networks…Dropout: a simple way to prevent neural networks from overfittingAttention Is All YouNeedAttention Is All You NeedModel-AgnosticMeta-Learning for Fast…Model-Agnostic Meta-Learning for Fast Adaptation of Deep NetworksLanguage Models asKnowledge Bases?Language Models as Knowledge Bases?BERT: Pre-training ofDeep Bidirectional…BERT: Pre-training of Deep Bidirectional Transformers for Language UnderstandingModifying Memories inTransformer ModelsModifying Memories in Transformer ModelsEditable Neural NetworksEditable Neural NetworksMultilingualAutoregressive Entity…Multilingual Autoregressive Entity LinkingGPT Understands, TooGPT Understands, TooOn the Opportunities andRisks of Foundation…On the Opportunities and Risks of Foundation ModelsFast Model Editing atScaleFast Model Editing at ScalePatching open-vocabularymodels by interpolating…Patching open-vocabulary models by interpolating weightsA Review on LanguageModels as Knowledge…A Review on Language Models as Knowledge BasesMeasuring andManipulating Knowledge…Measuring and Manipulating Knowledge Representations in Language ModelsEditing Models with TaskArithmeticEditing Models with Task ArithmeticPMET: Precise ModelEditing in a TransformerPMET: Precise Model Editing in a TransformerA Survey onHallucination in Large…A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open QuestionsWilKE: Wise-LayerKnowledge Editor for…WilKE: Wise-Layer Knowledge Editor for Lifelong Knowledge EditingPerturbation-RestrainedSequential Model EditingPerturbation-Restrained Sequential Model EditingMitigating HeterogeneousToken Overfitting in LL…Mitigating Heterogeneous Token Overfitting in LLM Knowledge EditingUnlocking Efficient,Scalable, and Continual…Unlocking Efficient, Scalable, and Continual Knowledge Editing with Basis-Level Representation Fine-TuningEditing FactualKnowledge in Language…Editing Factual Knowledge in Language Models過去の参考文献中心の論文この論文を引用する論文古い新しい

ノードをクリックするとフォーカスを固定、空白をクリックすると本論文に戻ります。ホバーで一時的にプレビューできます。各ノードのページはタイトルから開けます。