3 papers
cs.AI2026
Litmus: Zero-Label, Code-Driven Metric Specification for Evaluating AI Systems
Prajjwal Gupta, Prasang Gupta, Vishal Bhutani +4
As agentic LLM systems move from prototypes to deployment across increasingly diverse domains, evaluating them has become both more important and more difficult. The challenge is n…
cs.CL2025
LLMRank: Understanding LLM Strengths for Model Routing
Shubham Agrawal, Prasang Gupta
The rapid growth of large language models (LLMs) with diverse capabilities, latency and computational costs presents a critical deployment challenge: selecting the most suitable mo…
cs.AI2022
Intelligent Systematic Investment Agent: an ensemble of deep learning and evolutionary strategies
Prasang Gupta, Shaz Hoda, Anand Rao
Machine learning driven trading strategies have garnered a lot of interest over the past few years. There is, however, limited consensus on the ideal approach for the development o…