Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
Learning by Analogy: A Causal Framework for Composition Generalization
Lingjing Kong, Shaoan Xie, Yang Jiao +6
Compositional generalization -- the ability to understand and generate novel combinations of learned concepts -- enables models to extend their capabilities beyond limited experien…
cs.LG2025
Enabling Pareto-Stationarity Exploration in Multi-Objective Reinforcement Learning: A Multi-Objective Weighted-Chebyshev Actor-Critic Approach
Fnu Hairi, Jiao Yang, Tianchen Zhou +6
In many multi-objective reinforcement learning (MORL) applications, being able to systematically explore the Pareto-stationary solutions under multiple non-convex reward objectives…
cs.LG2023
Bandit Learning to Rank with Position-Based Click Models: Personalized and Equal Treatments
Tianchen Zhou, Jia Liu, Yang Jiao +4
Online learning to rank (ONL2R) is a foundational problem for recommender systems and has received increasing attention in recent years. Among the existing approaches for ONL2R, a…