2 papers
cs.LG2025
Benchmarking the Generality of Vision-Language-Action Models
Pranav Guruprasad, Sudipta Chowdhury, Harsh Sikka +6
Generalist multimodal agents are expected to unify perception, language, and control - operating robustly across diverse real world domains. However, current evaluation practices r…
cs.LG2025
Test-time augmentation improves efficiency in conformal prediction
Divya Shanmugam, Helen Lu, Swami Sankaranarayanan +1
A conformal classifier produces a set of predicted classes and provides a probabilistic guarantee that the set includes the true class. Unfortunately, it is often the case that con…