1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2024
Causal prompting model-based offline reinforcement learning
Xuehui Yu, Yi Guan, Rujia Shen +3
Model-based offline Reinforcement Learning (RL) allows agents to fully utilise pre-collected datasets without requiring additional or unethical explorations. However, applying mode…
cs.AI2024★ 1 cited
Dissociation of Faithful and Unfaithful Reasoning in LLMs
Evelyn Yee, Alice Li, Chenyu Tang +3
Large language models (LLMs) often improve their performance in downstream tasks when they generate Chain of Thought reasoning text before producing an answer. We investigate how L…