2 papers
cs.RO2026
Robot Critics that Sweat the Small Stuff
Sruthi Sudhakar, Junbang Liang, Sreehari Rammohan +3
Large vision-language models contain several priors about the world and object interactions, making them useful critics during inference to steer robot policies towards success. Ho…
cs.CV2026
: Smaller Self-Supervised ViTs Localize Better than Larger Ones
Sreehari Rammohan, Huy Ha, Carl Vondrick
Robust visual classification often depends on localizing the main foreground objects in an image while ignoring contextual distractors. Surprisingly, we find that the attention map…