2 papers
cs.CL2026
Constrained Semantic Decompression in LLMs through Persian Proverb-Conditioned Story Generation
Zahra Habibzadeh, Paria Khoshtab, Amir Mesbah +1
Transforming a dense, abstract proverb into an engaging and morally faithful narrative requires deep cultural understanding and robust semantic grounding. We frame this problem as…
cs.LG2025
Subgoal Discovery Using a Free Energy Paradigm and State Aggregations
Amirhossein Mesbah, Reshad Hosseini, Seyed Pooya Shariatpanahi +1
Reinforcement learning (RL) plays a major role in solving complex sequential decision-making tasks. Hierarchical and goal-conditioned RL are promising methods for dealing with two…