From the 1 of 4 linked papers with an AI index.
4 papers
Latent Fact-Checking: Detecting Misinformation through Activation Engineering
Pedro T. Barcelos, Pedro Barcelos, Otávio Parraga +4
The proliferation of misinformation online has driven demand for scalable detection systems. While most existing approaches rely on surface-level linguistic features or external kn…
Inference-Time Machine Unlearning via Gated Activation Redirection
VinÃcius Conte Turani, Otávio Parraga, João Vitor Boer Abitante +7
The paper proposes GUARD-IT, a gradient‑free method that modifies activations at inference time with input‑dependent rotations to erase specific data from large language models whi…
Low-Rank Adapters Initialization via Gradient Surgery for Continual Learning
Joana Pasquali, Ramiro N. Barros, Arthur S. Bianchessi +7
LoRA is widely adopted for continual fine-tuning of Large Language Models due to its parameter efficiency, modularity across tasks, and compatibility with replay strategies. Howeve…
Inference Time Debiasing Concepts in Diffusion Models
Lucas S. Kupssinskü, Marco N. Bochernitsan, Jordan Kopper +2
We propose DeCoDi, a debiasing procedure for text-to-image diffusion-based models that changes the inference procedure, does not significantly change image quality, has negligible…