Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
How Well Can Knowledge Edit Methods Edit Perplexing Knowledge?
Huaizhi Ge, Frank Rudzicz, Zining Zhu
Large language models (LLMs) have demonstrated remarkable capabilities, but updating their knowledge post-training remains a critical challenge. While recent model editing techniqu…
cs.CL2024
Understanding Language Model Circuits through Knowledge Editing
Huaizhi Ge, Frank Rudzicz, Zining Zhu
Recent advances in language model interpretability have identified circuits, critical subnetworks that replicate model behaviors, yet how knowledge is structured within these cruci…