From the 1 of 4 linked papers with an AI index.
4 papers
Not as Sweet by Another Name: An Empirical Study of Format Robustness in LLM Document Workflows
Xiaoyu Zhang, Xianyun Cheng, Tianlin Li +3
The paper investigates how changing document formats (e.g., CSV vs. plain text) affects the reliability of end-to-end LLM-driven software workflows, introducing a metamorphic testi…
Ensemble-Based Uncertainty Estimation for Code Correctness Estimation
Yunxiang Wei, Tianlin Li, Yuwei Zheng +6
Large language models (LLMs) have demonstrated remarkable capabilities in generating programs from natural language descriptions, yet ensuring their correctness without an external…
How Emotion Shapes the Behavior of LLMs and Agents: A Mechanistic Study
Moran Sun, Tianlin Li, Yuwei Zheng +4
Emotion plays an important role in human cognition and performance. Motivated by this, we investigate whether analogous emotional signals can shape the behavior of large language m…
Robust Multi-Agent Reinforcement Learning by Mutual Information Regularization
Simin Li, Ruixiao Xu, Jingqiao Xiu +4
In multi-agent reinforcement learning (MARL), ensuring robustness against unpredictable or worst-case actions by allies is crucial for real-world deployment. Existing robust MARL m…