From the 1 of 10 linked papers with an AI index.
10 papers
RoboDesign1M: A Large-scale Dataset for Robot Design Understanding
Tri Le, Toan Nguyen, Quang Tran +6
The paper presents RoboDesign1M, a million‑sample multimodal dataset of robot designs collected from scientific literature, and shows its usefulness for tasks such as design image…
ROAD-VLA: Robust Online Adaptation via Self-Distillation for Vision-Language-Action Models
Kejing Wang, Toan Nguyen, Minh Hoang Nguyen +2
Effective online adaptation of vision-language-action (VLA) models remains challenging, as sparse rewards provide weak supervision for high-dimensional autoregressive action polici…
CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning
Yang Liu, Toan Nguyen, Flora D. Salim
Catastrophic forgetting remains a major obstacle to continual learning in large language models (LLMs) and vision--language models (VLMs). Although Mixture-of-Experts (MoE) archite…
SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation
Taisei Hanyu, Nhat Chung, Huy Le +10
Inspired by how humans reason over discrete objects and their relationships, we explore whether compact object-centric and object-relation representations can form a foundation for…
ByteRover: Agent-Native Memory Through LLM-Curated Hierarchical Context
Andy Nguyen, Danh Doan, Hoang Pham +8
Memory-Augmented Generation (MAG) extends large language models with external memory to support long-context reasoning, but existing approaches universally treat memory as an exter…
HyperTokens: Controlling Token Dynamics for Continual Video-Language Understanding
Toan Nguyen, Yang Liu, Celso De Melo +1
Continual VideoQA with multimodal LLMs is hindered by interference between tasks and the prohibitive cost of storing task-specific prompts. We introduce HyperTokens, a transformer-…