2 papers
cs.LG2026
Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning
Keegan Harris, Brian W. Lee, Ian Waudby-Smith +3
Reinforcement learning (RL) fine-tuning is widely used in language model training to improve model performance on a target task while limiting drift from a reference policy. A stan…
cs.CV2025
Can ChatGPT Learn My Life From a Week of First-Person Video?
Keegan Harris
Motivated by recent improvements in generative AI and wearable camera devices (e.g. smart glasses and AI-enabled pins), I investigate the ability of foundation models to learn abou…