2 citations · 2 across the 2 of their papers we have counts for
4 papers
Inference-Time Scaling for Generalist Reward Modeling
Zijun Liu, Peiyi Wang, Runxin Xu +5
Reinforcement learning (RL) has been widely adopted in post-training for large language models (LLMs) at scale. Recently, the incentivization of reasoning capabilities in LLMs from…
Advancing Language Multi-Agent Learning with Credit Re-Assignment for Interactive Environment Generalization
Zhitao He, Zijun Liu, Peng Li +5
LLM-based agents have made significant advancements in interactive environments, such as mobile operations and web browsing, and other domains beyond computer using. Current multi-…
AIGS: Generating Science from AI-Powered Automated Falsification
Zijun Liu, Kaiming Liu, Yiqi Zhu +5
Rapid development of artificial intelligence has drastically accelerated the development of scientific discovery. Trained with large-scale observation data, deep neural networks ex…
Interactive Visual Assessment for Text-to-Image Generation Models
Xiaoyue Mi, Fan Tang, Juan Cao +5
Visual generation models have achieved remarkable progress in computer graphics applications but still face significant challenges in real-world deployment. Current assessment appr…