4 papers
Debate with Images: Detecting Deceptive Behaviors in Multimodal Large Language Models
Sitong Fang, Shiyi Hou, Kaile Wang +6
Are frontier AI systems becoming more capable? Certainly. Yet such progress is not an unalloyed blessing but rather a Trojan horse: behind their performance leaps lie more insidiou…
Conformal Agent Error Attribution
Naihe Feng, Yi Sui, Shiyi Hou +2
When multi-agent systems (MAS) fail, identifying where the decisive error occurred is the first step for automated recovery to an earlier state. Error attribution remains a fundame…
Textual Bayes: Quantifying Prompt Uncertainty in LLM-Based Systems
Brendan Leigh Ross, Noël Vouitsis, Atiyeh Ashari Ghomi +8
Although large language models (LLMs) are becoming increasingly capable of solving challenging real-world tasks, accurately quantifying their uncertainty remains a critical open pr…
Response Quality Assessment for Retrieval-Augmented Generation via Conditional Conformal Factuality
Naihe Feng, Yi Sui, Shiyi Hou +2
Existing research on Retrieval-Augmented Generation (RAG) primarily focuses on improving overall question-answering accuracy, often overlooking the quality of sub-claims within gen…