2 papers
cs.LG2024
IBCB: Efficient Inverse Batched Contextual Bandit for Behavioral Evolution History
Yi Xu, Weiran Shen, Xiao Zhang +1
Traditional imitation learning focuses on modeling the behavioral mechanisms of experts, which requires a large amount of interaction history generated by some fixed expert. Howeve…
cs.CL2024
On the Decision-Making Abilities in Role-Playing using Large Language Models
Chenglei Shen, Guofu Xie, Xiao Zhang +1
Large language models (LLMs) are now increasingly utilized for role-playing tasks, especially in impersonating domain-specific experts, primarily through role-playing prompts. When…