Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
FLARE: Few-shot Learning-based Adaptive Reflective Engine
Dhanasekar Sundararaman, Bharat Gandhi, Aashna Garg +1
Large language models (LLMs) are increasingly deployed in complex, compound AI systems where performance hinges on the quality of prompts. Recent state-of-the-art optimizers like G…
cs.CL2026
HyDRA: Hybrid Dynamic Routing Architecture for Heterogeneous LLM Pools
Aashna Garg, Siddharth Singha Roy, Jinu Jang +2
Production LLM deployments increasingly maintain heterogeneous model pools spanning order-of-magnitude cost differences. Existing routers make binary strong-vs-weak decisions and c…
cs.CL2025
LOCUS: A System and Method for Low-Cost Customization for Universal Specialization
Dhanasekar Sundararaman, Keying Li, Wayne Xiong +1
We present LOCUS (LOw-cost Customization for Universal Specialization), a pipeline that consumes few-shot data to streamline the construction and training of NLP models through tar…