works on

From the 1 of 11 linked papers with an AI index.

activity
20242026
collaborators

11 papers

cs.RO2026

Explicit Kinematic Guidance from Analytic Concepts for Vision-Language-Action Models

Mingyang Sun, Jiude Wei, Xiujian Liang +4

The paper introduces a Concept Expert module that extracts 3D kinematic and structural information to create Analytic Concepts, providing explicit guidance for Vision-Language-Acti…

cs.RO2026

Analytic Concept-Centric Memory for Agentic Embodied Manipulation

Mingyang Sun, Xiujian Liang, Jiude Wei +4

Long-horizon embodied manipulation requires agents to remember persistent objects, track changing scene states, and reuse prior interaction knowledge. However, existing agent memor…

cs.AI2026

OptArgus: A Multi-Agent System to Detect Hallucinations in LLM-based Optimization Modeling

Zhong Li, Zihan Guo, Xiaohan Lu +5

Large language models (LLMs) are increasingly used to translate natural-language optimization problems into mathematical formulations and solver code, but matching the reference ob…

cs.CV2025

SSR: Enhancing Depth Perception in Vision-Language Models via Rationale-Guided Spatial Reasoning

Yang Liu, Ming Ma, Xiaomin Yu +5

Despite impressive advancements in Visual-Language Models (VLMs) for multi-modal tasks, their reliance on RGB inputs limits precise spatial understanding. Existing methods for inte…

cs.LG2025

Iterative Refinement of Flow Policies in Probability Space for Online Reinforcement Learning

Mingyang Sun, Pengxiang Ding, Weinan Zhang +1

While behavior cloning with flow/diffusion policies excels at learning complex skills from demonstrations, it remains vulnerable to distributional shift, and standard RL methods st…

cs.RO2025

Executable Analytic Concepts as the Missing Link Between VLM Insight and Precise Manipulation

Mingyang Sun, Jiude Wei, Qichen He +3

Enabling robots to perform precise and generalized manipulation in unstructured environments remains a fundamental challenge in embodied AI. While Vision-Language Models (VLMs) hav…