2 papers
cs.LG2026
CLEANER: Self-Purified Trajectories Boost Agentic Reinforcement Learning
Tianshi Xu, Yuteng Chen, Meng Li
Agentic Reinforcement Learning (RL) has empowered Large Language Models (LLMs) to utilize tools like Python interpreters for complex problem-solving. However, for parameter-constra…
cs.CR2025
CryptoMoE: Privacy-Preserving and Scalable Mixture of Experts Inference via Balanced Expert Routing
Yifan Zhou, Tianshi Xu, Jue Hong +2
Private large language model (LLM) inference based on cryptographic primitives offers a promising path towards privacy-preserving deep learning. However, existing frameworks only s…