10 citations · 38 across the 31 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Beyond Confidence: Stability-Aware Test-Time Adaptation for LLM Reasoning
Bincheng Gu, Min Gao, Zongwei Wang +3
Test-time adaptation has emerged as a lightweight alternative to costly post-training for improving the reasoning capabilities of Large Language Models (LLMs) on downstream tasks.…
cs.AI2026
ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning
Jiaqi Zhang, Tong Chen, Junliang Yu +2
Agentic systems have rapidly advanced in their ability to interact with real-world environments, leverage external tools, and provide services for users. However, unlike natural-wo…
cs.AI2026
When Agents See Humans as the Outgroup: Belief-Dependent Bias in LLM-Powered Agents
Zongwei Wang, Bincheng Gu, Hongyu Yu +5
This paper reveals that LLM-powered agents exhibit not only demographic bias (e.g., gender, religion) but also intergroup bias under minimal "us" versus "them" cues. When such grou…