11 papers
Hybrid Retriever Evolution for Multimodal Document Reasoning Agents
Bohan Yao, Shruthan Radhakrishna, Vikas Yadav
Different retrievers, including lexical, semantic, and multimodal approaches, provide highly complementary strengths for multimodal document understanding, yet most systems combine…
Glass Box at Orbit: A Constitutional AI Verification Framework for Trustworthy Autonomous CubeSat Intelligence
Karthik Barma, Anil Sanneboyina, V C Premchand Yadav
The space industry is quietly building toward something nobody has fully reckoned with: orbital data centers running thousands of autonomous AI workloads with no human in the loop,…
ARM: Discovering Agentic Reasoning Modules for Generalizable Multi-Agent Systems
Bohan Yao, Shiva Krishna Reddy Malay, Vikas Yadav
Large Language Model (LLM)-powered Multi-agent systems (MAS) have achieved state-of-the-art results on various complex reasoning tasks. Recent works have proposed techniques to aut…
R2V Agent: Teaching SLMs When to Ask for Help
Raghu Vamshi Hemadri, Humaira Firdowse Mohammed, Rishabh Maheshwary +5
Efficient agentic systems should incur expensive frontier-model costs only on decisions where a cheaper local model is likely to fail. Existing LLM cascades usually route whole que…
AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs
Hoang Nguyen, Sidharth Surapaneni, Akshay Kalkunte +9
Large Audio Language Models (LALMs) are rapidly advancing, but evaluating them remains challenging due to inefficient and non-standardized toolkits that limit fair comparison and s…
Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems
Nishanth Madhusudhan, Vikas Yadav, Alexandre Lacoste
Effective abstention (EA), recognizing evidence insufficiency and refraining from answering, is critical for reliable multimodal systems. Yet existing evaluation paradigms for visi…