collaborators

11 papers

cs.LG2026

Observational Policy Ranking for SMB Financial Guidance from Multi-Action Accounting Logs

Shrutendra Harsola, Vignesh Subrahmaniam, Vikas Raturi +7

Small and medium-sized businesses need timely financial guidance, yet historical accounting logs record self-selected and often co-occurring business changes rather than randomized…

cs.CL2026

RIMRULE: Improving Tool-Using Language Agents via MDL-Guided Rule Learning

Xiang Gao, Yuguang Yao, Qi Zhang +5

Large language models (LLMs) often struggle to use tools reliably in domain-specific settings, where APIs may be idiosyncratic, under-documented, or tailored to private workflows.…

cs.LG2026

Textual Belief States for World Models: Identifiable Representation Learning Under Strict Mediation

Xiang Gao, Kaiwen Dong, Yuguang Yao +2

World models in partially observed environments rely on latent representations that summarize interaction history, but in many modern LLM-based architectures predictive performance…

cs.CL2026

Executable Schema Contracts: From Automatic Ingestion to Multi-Source Retrieval

Padmaja Jonnalagedda, Yuguang Yao, Xiang Gao +2

Real-world data spans tables, documents, and semi-structured files with implicit semantics. Querying this data requires integrating evidence across inconsistent schemas and formats…

cs.CL2026

When Evidence is Sparse: Weakly Supervised Early Failure Alerting in Dialogs and LLM-Agent Trajectories

Avinash Baidya, Xinran Liang, Ruocheng Guo +2

Early failure alerting requires deciding, while a dialog or agent trajectory is still unfolding, whether to flag it as likely to fail. This is challenging because supervision is ty…

cs.AI2026

Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use

Ruocheng Guo, Kaiwen Dong, Xiang Gao +1

While most efforts to improve LLM-based tool-using agents focus on the agent itself - through larger models, better prompting, or fine-tuning - agent performance increasingly plate…