3 papers
cs.LG2026
Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training
Xi Wang, Wenbo Lu, Shengjie Wang
Generative Flow Networks (GFlowNets) enable fine-tuning large language models to approximate reward-proportional posteriors, but they remain prone to mode collapse, manifesting as…
cs.CL2026
Eval4Sim: An Evaluation Framework for Persona Simulation
Eliseo Bao, Anxo Perez, Xi Wang +1
Large Language Model personas, explicit profiles specifying a user's attributes, preferences, and behavioural tendencies, are increasingly used to simulate human conversations for…
cs.CL2025
TalkDep: Clinically Grounded LLM Personas for Conversation-Centric Depression Screening
Xi Wang, Anxo Perez, Javier Parapar +1
The increasing demand for mental health services has outpaced the availability of real training data to develop clinical professionals, leading to limited support for the diagnosis…