2 papers
cs.HC2026
Navigating the Conceptual Multiverse
Andre Ye, Jenny Y. Huang, Alicia Guo +3
When language models answer open-ended problems, they implicitly make hidden decisions that shape their outputs, leaving users with uncontextualized answers rather than a working m…
cs.LG2025
Latency and Token-Aware Test-Time Compute
Jenny Y. Huang, Mehul Damani, Yousef El-Kurdi +2
Inference-time scaling has emerged as a powerful way to improve large language model (LLM) performance by generating multiple candidate responses and selecting among them. However,…