works on

From the 1 of 5 linked papers with an AI index.

collaborators

5 papers

physics.chem-ph2026

MDArena: Evaluating Coding Agents on Realistic Molecular Dynamics Workflows

Nithishwer Mouroug Anand, Wei-Tse Hsu, Kyle Vaccaro +6

Accelerating scientific discovery is among the most consequential applications of AI, and computational biomolecular simulation stands out as a particularly promising target within…

cs.RO2026

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation

Byeongguk Jeon, Seonghyeon Ye, JaeHyeok Doo +4

RoboWorld is an automated pipeline that uses a fast autoregressive video world model and a vision-language scoring system to evaluate generalist robot policies efficiently and reli…

cs.LG2026

Q-Flow: Stable and Expressive Reinforcement Learning with Flow-Based Policy

JaeHyeok Doo, Byeongguk Jeon, Seonghyeon Ye +2

There is growing interest in utilizing flow-based models as decision-making policies in reinforcement learning due to their high expressive capacity. However, effectively leveragin…

cs.CL2026

Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressive Diffusion Language Models

Jiyeon Kim, Sungik Choi, Yongrae Jo +2

Diffusion-based language models (dLLMs) have emerged as a promising alternative to autoregressive language models, offering the potential for parallel token generation and bidirect…

cs.CL2026

Can Large Language Models Keep Up? Benchmarking Online Adaptation to Continual Knowledge Streams

Jiyeon Kim, Hyunji Lee, Dylan Zhou +6

LLMs operating in dynamic real-world contexts often encounter knowledge that evolves continuously or emerges incrementally. To remain accurate and effective, models must adapt to n…