2 citations · 2 across the 2 of their papers we have counts for
4 papers
CycliST: A Video Language Model Benchmark for Reasoning on Cyclical State Transitions
Simon Kohaut, Daniel Ochs, Shun Zhang +4
We present CycliST, a novel benchmark dataset designed to evaluate Video Language Models (VLM) on their ability for textual reasoning over cyclical state transitions. CycliST captu…
GundamQ: Multi-Scale Spatio-Temporal Representation Learning for Robust Robot Path Planning
Yutong Shen, Ruizhe Xia, Bokai Yan +4
In dynamic and uncertain environments, robotic path planning demands accurate spatiotemporal environment understanding combined with robust decision-making under partial observabil…
Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble
Shun Zhang, Zhenfang Chen, Sunli Chen +3
Reinforcement Learning from Human Feedback (RLHF) is a widely adopted approach for aligning large language models with human values. However, RLHF relies on a reward model that is…
Adaptive Online Replanning with Diffusion Models
Siyuan Zhou, Yilun Du, Shun Zhang +5
Diffusion models have risen as a promising approach to data-driven planning, and have demonstrated impressive robotic control, reinforcement learning, and video planning performanc…