3 papers
cs.AI2024
WILT: A Multi-Turn, Memorization-Robust Inductive Logic Benchmark for LLMs
Eryk Banatt, Jonathan Cheng, Skanda Vaidyanath +1
While large language models have shown impressive capabilities across a wide range of domains, they still encounter significant challenges in reasoning tasks that require gathering…
cs.LG2023
Hindsight-DICE: Stable Credit Assignment for Deep Reinforcement Learning
Akash Velu, Skanda Vaidyanath, Dilip Arumugam
Oftentimes, environments for sequential decision-making problems can be quite sparse in the provision of evaluative feedback to guide reinforcement-learning agents. In the extreme…
cs.CV2023
Differentiable Weight Masks for Domain Transfer
Samar Khanna, Skanda Vaidyanath, Akash Velu
One of the major drawbacks of deep learning models for computer vision has been their inability to retain multiple sources of information in a modular fashion. For instance, given…