1 citations · 1 across the 3 of their papers we have counts for
3 papers
Cognify: Supercharging Gen-AI Workflows With Hierarchical Autotuning
Zijian He, Reyna Abhyankar, Vikranth Srivatsa +1
Today's gen-AI workflows that involve multiple ML model calls, tool/API calls, data retrieval, or generic code execution are often tuned manually in an ad-hoc way that is both time…
Preble: Efficient Distributed Prompt Scheduling for LLM Serving
Vikranth Srivatsa, Zijian He, Reyna Abhyankar +2
Prompts to large language models (LLMs) have evolved beyond simple user questions. For LLMs to solve complex problems, today's practices are to include domain-specific instructions…
The Effect of Model Size on Worst-Group Generalization
Alan Pham, Eunice Chan, Vikranth Srivatsa +6
Overparameterization is shown to result in poor test accuracy on rare subgroups under a variety of settings where subgroup information is known. To gain a more complete picture, we…