works on

From the 2 of 10 linked papers with an AI index.

collaborators

10 papers

cs.RO2026

Trajectory Divergence Horizon Decision for Reliable Dual-Arm Surgical Subtask Manipulation

Mingwu Su, Guankun Wang, Jinsong Lin +8

Surgical robotic systems are increasingly being adopted as clinical workload rises, motivating autonomous solutions for repetitive manipulation subtasks. Learning-based controllers…

cs.CL2026

SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning

Jinyang Wu, Shuo Yang, Zhengxi Lu +8

The paper introduces SEED, a framework that extracts reusable natural-language skills from on-policy trajectories and distills them back into the policy to provide dense token-leve…

cs.CL2026

DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment

Xinyu Geng, Xuanhua He, Sixiang Chen +7

The paper introduces DeepSearch-World, a deterministic, verifiable web environment, and DeepSearch-Evolve, a self‑distillation framework that lets web search agents improve from th…

cs.LG2026

Embedding-Based Federated Learning with Runtime Governance for Iron Deficiency Prediction

Fan Zhang, Simon Deltadahl, Majid Lotfian Delouee +10

Recent reviews find that the vast majority of published healthcare federated learning (FL) studies never reach real-world deployment. We developed an embedding-based FL pipeline fo…

cs.AI2026

Evaluating the Utility of Personal Health Records in Personalized Health AI

Rory Sayres, Kejia Chen, Ayush Jain +19

Patient-managed Personal Health Records (PHRs) promises to empower patients to better understand their health; but information in the record is complex, potentially hindering insig…

cs.LG2026

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles

Jinyang Wu, Guocheng Zhai, Ruihan Jin +7

The proliferation of large language models (LLMs) and modular skills has endowed autonomous agents with increasingly powerful capabilities. Existing frameworks typically rely on mo…