1 citations · 1 across the 2 of their papers we have counts for
4 papers
Not All Errors Are Created Equal: ASCoT Addresses Late-Stage Fragility in Efficient LLM Reasoning
Dongxu Zhang, Yujun Wu, Yiding Sun +5
While Chain-of-Thought (CoT) prompting empowers Large Language Models (LLMs), ensuring reasoning reliability remains an open challenge. Contrary to the prevailing cascading failure…
Uncertainty-aware Reward Design Process
Yang Yang, Xiaolu Zhou, Bosong Ding +1
Designing effective reward functions is a cornerstone of reinforcement learning (RL), yet it remains a challenging process due to the inefficiencies and inconsistencies inherent in…
BFRFormer: Transformer-based generator for Real-World Blind Face Restoration
Guojing Ge, Qi Song, Guibo Zhu +5
Blind face restoration is a challenging task due to the unknown and complex degradation. Although face prior-based methods and reference-based methods have recently demonstrated hi…
We Choose to Go to Space: Agent-driven Human and Multi-Robot Collaboration in Microgravity
Miao Xin, Zhongrui You, Zihan Zhang +7
We present SpaceAgents-1, a system for learning human and multi-robot collaboration (HMRC) strategies under microgravity conditions. Future space exploration requires humans to wor…