collaborators

5 papers

cs.AI2026

MemFail: Stress-Testing Failure Modes of LLM Memory Systems

Ishir Garg, Neel Kolhe, Dawn Song +1

Large language model (LLM) agents increasingly rely on external memory systems to remain consistent across long-horizon interactions, but little empirical work has been done to und…

cs.CL2026

InfoSynth: Information-Guided Benchmark Synthesis for LLMs

Ishir Garg, Neel Kolhe, Xuandong Zhao +1

Large language models (LLMs) have demonstrated significant advancements in reasoning and code generation, but efficiently creating new benchmarks to evaluate these capabilities rem…

cs.CL2026

Reliable Fine-Grained Evaluation of Natural Language Math Proofs

Wenjie Ma, Andrei Cojocaru, Neel Kolhe +6

Recent advances in large language models (LLMs) for mathematical reasoning have largely focused on tasks with easily verifiable final answers while generating and verifying natural…

astro-ph.GA2026

WLM: Dynamics of an isolated Dwarf Irregular Galaxy Under Ram Pressure in the Local Group

Neel Kolhe, Francois Hammer, Yanbin Yang +5

WLM is an archetypal dwarf irregular galaxy that has not experienced interactions with major Local Group galaxies within the past 8 Gyr. It has recently been shown that WLM is losi…

cs.LG2026

Fisher-Orthogonal Projected Natural Gradient Descent for Continual Learning

Ishir Garg, Neel Kolhe, Andy Peng +1

Continual learning aims to enable neural networks to acquire new knowledge on sequential tasks. However, the key challenge in such settings is to learn new tasks without catastroph…