From the 1 of 5 linked papers with an AI index.
5 papers
MDArena: Evaluating Coding Agents on Realistic Molecular Dynamics Workflows
Nithishwer Mouroug Anand, Wei-Tse Hsu, Kyle Vaccaro +6
Accelerating scientific discovery is among the most consequential applications of AI, and computational biomolecular simulation stands out as a particularly promising target within…
RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation
Byeongguk Jeon, Seonghyeon Ye, JaeHyeok Doo +4
RoboWorld is an automated pipeline that uses a fast autoregressive video world model and a vision-language scoring system to evaluate generalist robot policies efficiently and reli…
Q-Flow: Stable and Expressive Reinforcement Learning with Flow-Based Policy
JaeHyeok Doo, Byeongguk Jeon, Seonghyeon Ye +2
There is growing interest in utilizing flow-based models as decision-making policies in reinforcement learning due to their high expressive capacity. However, effectively leveragin…
Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressive Diffusion Language Models
Jiyeon Kim, Sungik Choi, Yongrae Jo +2
Diffusion-based language models (dLLMs) have emerged as a promising alternative to autoregressive language models, offering the potential for parallel token generation and bidirect…
Can Large Language Models Keep Up? Benchmarking Online Adaptation to Continual Knowledge Streams
Jiyeon Kim, Hyunji Lee, Dylan Zhou +6
LLMs operating in dynamic real-world contexts often encounter knowledge that evolves continuously or emerges incrementally. To remain accurate and effective, models must adapt to n…