1 citations · 1 across the 3 of their papers we have counts for
8 papers
Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning
Rufeng Chen, Zhaofan Zhang, Zhejiang Yang +2
Offline safe reinforcement learning often requires policies to adapt at deployment time to safety budgets that vary across episodes or change within a single episode. While diffusi…
TGSBM: Transformer-Guided Stochastic Block Model for Link Prediction
Zhejian Yang, Songwei Zhao, Zilin Zhao +1
Link prediction is a cornerstone of the Web ecosystem, powering applications from recommendation and search to knowledge graph completion and collaboration forecasting. However, la…
Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents
Zhejian Yang, Yongchao Chen, Xueyang Zhou +8
Long-horizon robotic manipulation poses significant challenges for autonomous systems, requiring extended reasoning, precise execution, and robust error recovery across complex seq…
Analytic Energy-Guided Policy Optimization for Offline Reinforcement Learning
Jifeng Hu, Sili Huang, Zhejian Yang +6
Conditional decision generation with diffusion models has shown powerful competitiveness in reinforcement learning (RL). Recent studies reveal the relation between energy-function-…
Towards Robust Multi-UAV Collaboration: MARL with Noise-Resilient Communication and Attention Mechanisms
Zilin Zhao, Chishui Chen, Haotian Shi +4
Efficient path planning for unmanned aerial vehicles (UAVs) is crucial in remote sensing and information collection. As task scales expand, the cooperative deployment of multiple U…
A Survey on Post-training of Large Language Models
Guiyao Tie, Zeli Zhao, Dingjie Song +23
The emergence of Large Language Models (LLMs) has fundamentally transformed natural language processing, making them indispensable across domains ranging from conversational system…