1 citations · 1 across the 1 of their papers we have counts for
9 papers
Policy Optimization and Statistical Inference for Online Contextual Matrix Games
Liner Xiang, Yixin Wang, Hengrui Cai
Online decision making often requires navigating a landscape shaped by both dynamic contexts and strategic interactions. In competitive pricing, for example, hotels must account fo…
Enhancing Causal Reasoning in Large Language Models: A Causal Attribution Model for Precision Fine-Tuning
Hengrui Cai, Shengjie Liu, Rui Song
This paper introduces a causal attribution model to enhance the interpretability of large language models (LLMs) and improve their causal reasoning abilities via precise fine-tunin…
PABU: Progress-Aware Belief Update for Efficient LLM Agents
Haitao Jiang, Lin Ge, Hengrui Cai +1
Large Language Model (LLM) agents commonly condition actions on full action-observation histories, which introduce task-irrelevant information that easily leads to redundant action…
Correct Reasoning Paths Visit Shared Decision Pivots
Dongkyu Cho, Amy B. Z. Zhang, Bilel Fehri +4
Chain-of-thought (CoT) reasoning exposes the intermediate thinking process of large language models (LLMs), yet verifying those traces at scale remains unsolved. In response, we in…
Text Rationalization for Robust Causal Effect Estimation
Lijinghua Zhang, Hengrui Cai
Recent advances in natural language processing have enabled the increasing use of text data in causal inference, particularly for adjusting confounding factors in treatment effect…
Foresighted Online Policy Optimization with Interference
Liner Xiang, Jiayi Wang, Hengrui Cai
Contextual bandits, which leverage the baseline features of sequentially arriving individuals to optimize cumulative rewards while balancing exploration and exploitation, are criti…