Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Stress Testing Unlearning Algorithms
Noam Diamant, Ethan Fetaya, Neta Glazer
Recently, machine unlearning, the removal of specific training data influence from a model, has gained increasing attention. In large language models (LLMs), unlearning is particul…
cs.LG2025
Multi Task Inverse Reinforcement Learning for Common Sense Reward
Neta Glazer, Aviv Navon, Aviv Shamsian +1
One of the challenges in applying reinforcement learning in a complex real-world environment lies in providing the agent with a sufficiently detailed reward function. Any misalignm…