Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
AdaptiveStep: Automatically Dividing Reasoning Step through Model Confidence
Yuliang Liu, Junjie Lu, Zhaoling Chen +10
Current approaches for training Process Reward Models (PRMs) often involve breaking down responses into multiple reasoning steps using rule-based techniques, such as using predefin…
cs.AI2024
Empowering Large Language Models on Robotic Manipulation with Affordance Prompting
Guangran Cheng, Chuheng Zhang, Wenzhe Cai +3
While large language models (LLMs) are successful in completing various language processing tasks, they easily fail to interact with the physical world by generating control sequen…