5 papers
SpanUQ: Span-Level Uncertainty Quantification for Large Language Model Generation
Yimeng Zhang, Yingying Zhuang, Ziyi Wang +12
Uncertainty estimation is essential not only for the trustworthy deployment of large language models (LLMs) but also as a foundation for self-refinement in LLM generation. However,…
TRIM: Hybrid Inference via Targeted Stepwise Routing in Multi-Step Reasoning Tasks
Vansh Kapoor, Aman Gupta, Hao Chen +3
Multi-step reasoning tasks like mathematical problem solving are vulnerable to cascading failures, where a single incorrect step leads to complete solution breakdown. Current LLM r…
VAL-Bench: Belief Consistency as a measure for Value Alignment in Language Models
Aman Gupta, Denny O'Shea, Fazl Barez
Large language models (LLMs) are increasingly being used for tasks where outputs shape human decisions, so it is critical to verify that their responses consistently reflect desire…
How and Where to Translate? The Impact of Translation Strategies in Cross-lingual LLM Prompting
Aman Gupta, Yingying Zhuang, Zhou Yu +2
Despite advances in the multilingual capabilities of Large Language Models (LLMs), their performance varies substantially across different languages and tasks. In multilingual retr…
Multilingual Information Retrieval with a Monolingual Knowledge Base
Yingying Zhuang, Aman Gupta, Anurag Beniwal
Multilingual information retrieval has emerged as powerful tools for expanding knowledge sharing across languages. On the other hand, resources on high quality knowledge base are o…