3 papers
cs.LG2025
Intra-Trajectory Consistency for Reward Modeling
Chaoyang Zhou, Shunyu Liu, Zengmao Wang +4
Reward models are critical for improving large language models (LLMs), particularly in reinforcement learning from human feedback (RLHF) or inference-time verification. Current rew…
cs.CV2025
TiMo: Spatiotemporal Foundation Model for Satellite Image Time Series
Xiaolei Qin, Di Wang, Jing Zhang +4
Satellite image time series (SITS) provide continuous observations of the Earth's surface, making them essential for applications such as environmental management and disaster asse…
cs.CV2025
DGSolver: Diffusion Generalist Solver with Universal Posterior Sampling for Image Restoration
Hebaixu Wang, Jing Zhang, Haonan Guo +3
Diffusion models have achieved remarkable progress in universal image restoration. While existing methods speed up inference by reducing sampling steps, substantial step intervals…