From the 1 of 5 linked papers with an AI index.
5 papers
RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation
Byeongguk Jeon, Seonghyeon Ye, JaeHyeok Doo +4
RoboWorld is an automated pipeline that uses a fast autoregressive video world model and a vision-language scoring system to evaluate generalist robot policies efficiently and reli…
HABIT: Human-Aware Behavior and Interaction Training Dataset for Robot Manipulation
Jaehwi Song, Suchae Jeong, Byeongguk Jeon +4
Large-scale demonstration datasets have been central to recent progress in general-purpose robot policies. However, existing datasets are collected in human-absent settings, and po…
Ask Optimal Questions: Aligning Large Language Models with Retriever's Preference in Conversation
Chanwoong Yoon, Gangwoo Kim, Byeongguk Jeon +3
Conversational search, unlike single-turn retrieval tasks, requires understanding the current question within a dialogue context. The common approach of rewrite-then-retrieve aims…
Generative Prompt Internalization
Haebin Shin, Lei Ji, Yeyun Gong +3
Prompts used in recent large language model based applications are often fixed and lengthy, leading to significant computational overhead. To address this challenge, we propose Gen…
Rethinking the Role of Proxy Rewards in Language Model Alignment
Sungdong Kim, Minjoon Seo
Learning from human feedback via proxy reward modeling has been studied to align Large Language Models (LLMs) with human values. However, achieving reliable training through that p…