8 papers
RoboLineage: Agent-Native Data Lifecycle Governance Across Robot Policy Iterations
Qian Luo, Wentao Guo, Zhennan Qin +4
We present RoboLineage, an agent-native data lifecycle governance system for robot policy iteration. Modern robot policies improve through repeated data collection, review, retrain…
Tri-Info: Generalizable, Interpretable Failure Prediction for VLA Models via Information Theory
Jinghan Yang, Yunchao Zhang, Wang Yuan +4
Vision-Language-Action (VLA) models are increasingly deployed across diverse tasks, yet they remain black boxes whose physical interactions can cause irreversible harm, making gene…
Distributed Image Compression with Multimodal Side Information at Extremely Low Bitrates
Guojun Xu, Mingyang Zhang, Jianwen Xiang +3
Distributed Image Compression (DIC) is crucial for multi-view transmission, especially when operating at extremely low bitrates (< 0.1 bpp). Its core challenge is effectively utili…
DISC: Decoupling Instruction from State-Conditioned Control via Policy Generation
Hanxiang Ren, Pei Zhou, Xunzhe Zhou +1
Language-conditioned manipulation policies typically process instructions and observations through shared network parameters. This task-state entanglement provides a pathway for ob…
DexHoldem: Playing Texas Hold'em with Dexterous Embodied System
Feng Chen, Tianzhe Chu, Li Sun +6
Evaluating embodied systems on real dexterous hardware requires more than isolated primitive skills: an agent must perceive a changing tabletop scene, choose a context-appropriate…
Gaze-Regularized Vision-Language-Action Models for Robotic Manipulation
Anupam Pani, Yanchao Yang
Despite advances in Vision-Language-Action (VLA) models, robotic manipulation struggles with fine-grained tasks because current models lack mechanisms for active visual attention a…