2 papers
cs.LG2026
Conformal Cascade: Distribution-Free Accuracy Guarantees for Multi-Tier LLM Inference
Yifan Dou, Shikan Lian, Shikan Fang +1
Large language model (LLM) cascades reduce inference cost by routing easy queries to a small model and deferring hard queries to a larger one. Production cascades govern this defer…
cs.LG2026
Verbalized Particle Posterior: Bayesian Inference over Natural Language Hypotheses
Yan Zhang, Shikan Lian, Shibo Li
Verbalized Machine Learning (VML) parameterizes a model as a natural-language prompt that an LLM evaluates as f(x; theta). The framework is interpretable, but it commits to a singl…