5 papers
Iterative Visual Thinking and the Self-Correction Mirage in VLM Grounding
Animesh Tripathy, Aswanth Krishnan
Letting a vision-language model (VLM) think longer at test time has driven much recent progress. A natural way to bring this to spatial grounding is visual self-correction: the mod…
Self-Specializing Vision-Language Transmon Chip Calibration in a Physics-Grounded Environment
Animesh Tripathy, Aswanth Krishnan
Calibrating a superconducting transmon chip is a sequential decision problem under noise, drift, and a finite budget: an expert must choose experiments, read ambiguous plots, judge…
Beyond the Performance Illusion: Structure-Aware Stratified Partitioning and Curriculum Distributionally Robust Optimization for Spatially Correlated Domains
Prathamesh Patil, Arpit Jain, Aswanth Krishnan
Performance evaluation in AI systems commonly assumes that random dataset splits produce independent and identically distributed (i.i.d.) subsets. We show that this assumption ofte…
GitOfThoughts: Version-Controlled Reasoning and Agent Memory You Can Replay, Diff, and Merge
Pavan C Shekar, Abhishek H S, Aswanth Krishnan
Large language model reasoning leaves no trace once it is done. The steps of a chain of thought disappear when the context window closes, a pruned search branch is just gone, and m…
A Benchmark for Procedural Memory Retrieval in Language Agents
Ishant Kohar, Aswanth Krishnan
Current AI agents excel in familiar settings, but fail sharply when faced with novel tasks with unseen vocabularies -- a core limitation of procedural memory systems. We present th…