1 citations · 1 across the 3 of their papers we have counts for
4 papers
AgentV-RL: Scaling Reward Modeling with Agentic Verifier
Jiazheng Zhang, Ziche Fu, Zhiheng Xi +13
Verifiers have been demonstrated to enhance LLM reasoning via test-time scaling (TTS). Yet, they face significant challenges in complex domains. Error propagation from incorrect in…
GUIDE: A Guideline-Guided Dataset for Instructional Video Comprehension
Jiafeng Liang, Shixin Jiang, Zekun Wang +7
There are substantial instructional videos on the Internet, which provide us tutorials for completing various tasks. Existing instructional video datasets only focus on specific st…
KwaiAgents: Generalized Information-seeking Agent System with Large Language Models
Haojie Pan, Zepeng Zhai, Hao Yuan +5
Driven by curiosity, humans have continually sought to explore and understand the world around them, leading to the invention of various tools to satiate this inquisitiveness. Desp…
CogGPT: Unleashing the Power of Cognitive Dynamics on Large Language Models
Yaojia Lv, Haojie Pan, Zekun Wang +6
Cognitive dynamics are pivotal to advance human understanding of the world. Recent advancements in large language models (LLMs) reveal their potential for cognitive simulation. How…