2 papers
cs.RO2026
RLDX-1 Technical Report
Dongyoung Kim, Huiwon Jang, Myungkyu Koo +65
While Vision-Language-Action models (VLAs) have shown remarkable progress toward human-like generalist robotic policies through the versatile intelligence (i.e. broad scene underst…
cs.HC2026
AgentLens: Adaptive Visual Modalities for Human-Agent Interaction in Mobile GUI Agents
Jeonghyeon Kim, Byeongjun Joung, Junwon Lee +3
Mobile GUI agents can automate smartphone tasks by interacting directly with app interfaces, but how they should communicate with users during execution remains underexplored. Exis…