Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Rational Tuning of LLM Cascades via Probabilistic Modeling
Michael J. Zellinger, Matt Thomson
Understanding the reliability of large language models (LLMs) has recently garnered significant attention. Given LLMs' propensity to hallucinate, as well as their high sensitivity…
cs.LG2024
Efficiently Deploying LLMs with Controlled Risk
Michael J. Zellinger, Matt Thomson
Deploying large language models in production requires simultaneous attention to efficiency and risk control. Prior work has shown the possibility to cut costs while maintaining si…