Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
TW-CRL: Time-Weighted Contrastive Reward Learning for Efficient Inverse Reinforcement Learning
Yuxuan Li, Yicheng Gao, Ning Yang +1
Episodic tasks in Reinforcement Learning (RL) often pose challenges due to sparse reward signals and high-dimensional state spaces, which hinder efficient learning. Additionally, t…
cs.LG2024
ElastiFormer: Learned Redundancy Reduction in Transformer via Self-Distillation
Junzhang Liu, Tingkai Liu, Yueyuan Sui +1
We introduce ElastiFormer, a post-training technique that adapts pretrained Transformer models into an elastic counterpart with variable inference time compute. ElastiFormer introd…