language model agents 1policy optimization 1reinforcement learning 1sandbox environments 1variance reduction 1
From the 1 of 8 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Pedagogically-Inspired Data Synthesis for Language Model Knowledge Distillation
Bowei He, Yankai Chen, Xiaokun Zhang +4
Knowledge distillation from Large Language Models (LLMs) to smaller models has emerged as a critical technique for deploying efficient AI systems. However, current methods for dist…
cs.AI2025
Embracing Trustworthy Brain-Agent Collaboration as Paradigm Extension for Intelligent Assistive Technologies
Yankai Chen, Xinni Zhang, Yifei Zhang +6
Brain-Computer Interfaces (BCIs) offer a direct communication pathway between the human brain and external devices, holding significant promise for individuals with severe neurolog…