Skip to main content
Aggregate arXiv cs.AI 人工智能 31 Aug 2026 - 15:00

KLOD: Locality-Preserving Knowledge Editing via Non-Target Distribution Preservation

RSS 官方收录 · 可信分层展示

关键摘要

arXiv:2608.…

  • 27839v1 Announce Type: new Abstract: Fine-tuning-based knowledge editi…
  • In sequential editing, such unconstrained redistribution can accumulat…
  • We propose KLOD, a bounded and distribution-preserving objective for f…

摘要引擎:抽取

正文提要

arXiv:2608.27839v1 Announce Type: new Abstract: Fine-tuning-based knowledge editing is simple and architecture-agnostic, but standard cross-entropy increases the edited target probability without explicitly constraining changes in the non-target output distribution. In sequential editing, such unconstrained redistribution can accumulate as distributional drift and contribute to locality degradation. We propose KLOD, a bounded and distribution-preserving objective for fine-tuning-based knowledge editing that separates the intended target update from distributions that should remain stable. KLOD stops target amplification once a probability threshold is reached, while preserving the target-excluded non-target distribution at target positions and the full next-token distribution at prefix positions. Experiments on CounterFact and ZsRE with Llama3-8B-Instruct and Qwen2.5-7B-Instruct show that KLOD substantially mitigates locality degradation while maintaining high edit reliability. The target probability threshold further provides a controllable Generalization--Locality trade-off. Ablation, multi-seed, and distributional KL analyses support the interpretation that KLOD's locality gains are associated with preserving output distributions rather than simply weakening the edit. Code is available on GitHub https://github.com/Hostoday/KLOD .

来源:https://arxiv.org/abs/2608.27839

打开官方原文 站点原文页 可信分区 本信源更多 今日简报 分享图 RSS 稍后再看列表