works on

From the 1 of 5 linked papers with an AI index.

activity
20242026
collaborators

5 papers

cs.RO2026

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation

Byeongguk Jeon, Seonghyeon Ye, JaeHyeok Doo +4

RoboWorld is an automated pipeline that uses a fast autoregressive video world model and a vision-language scoring system to evaluate generalist robot policies efficiently and reli…

cs.RO2026

HABIT: Human-Aware Behavior and Interaction Training Dataset for Robot Manipulation

Jaehwi Song, Suchae Jeong, Byeongguk Jeon +4

Large-scale demonstration datasets have been central to recent progress in general-purpose robot policies. However, existing datasets are collected in human-absent settings, and po…

cs.IR2025

Ask Optimal Questions: Aligning Large Language Models with Retriever's Preference in Conversation

Chanwoong Yoon, Gangwoo Kim, Byeongguk Jeon +3

Conversational search, unlike single-turn retrieval tasks, requires understanding the current question within a dialogue context. The common approach of rewrite-then-retrieve aims…

cs.CL2025

Generative Prompt Internalization

Haebin Shin, Lei Ji, Yeyun Gong +3

Prompts used in recent large language model based applications are often fixed and lengthy, leading to significant computational overhead. To address this challenge, we propose Gen…

cs.LG2024

Rethinking the Role of Proxy Rewards in Language Model Alignment

Sungdong Kim, Minjoon Seo

Learning from human feedback via proxy reward modeling has been studied to align Large Language Models (LLMs) with human values. However, achieving reliable training through that p…