activity
20242026
collaborators

12 papers

cs.DC2026

Energy-Efficient Multimodal Inference Serving with Tri-serve

Ziyang Jia, Sara Rashidi Golrouye, Laxmi Bhuyan +5

Multimodal model inference creates substantial energy demand with growing performance requirements. Within GPUs, power is autonomously managed by an on-board power management unit…

eess.SY2026

TetraRL: A Self-Adaptive Runtime for On-Device Deep Reinforcement Learning Systems

Zexin Li, Soheil Shirvani, Cong Liu

Autonomous robotic systems, including autonomous vehicles, drones, and mobile robots, increasingly rely on on-device Deep Reinforcement Learning (DRL) to adapt to dynamic environme…

eess.SY2026

Orion: Enabling Self-adaptive Memory Management for On-device Online Continual Learning

Zexin Li, Nikil Dutt, Cong Liu

Online continual learning (OCL) enables real-time adaptation to new data, making it crucial for dynamic robotic applications. However, its practical deployment is hindered by memor…

cs.RO2026

RED: Adaptive Real-Time DAG Scheduling for Robotic Inference under Environmental Dynamics

Zexin Li, Tao Ren, Johnathan Liu +2

Robots deployed in dynamic environments must contend with environment-driven changes that reshape computation at runtime: new tasks may appear, precedence relations can shift, and…

cs.RO2026

PIMbot: A Self-Adaptive Attack Framework for Adversarial Manipulation of Multi-Robot Reinforcement Learning

Zexin Li, Ziliang Zhang, Hyoseung Kim +1

Recent research has demonstrated the potential of reinforcement learning in effective multi-robot collaboration, particularly in social dilemmas where robots face a trade-off betwe…

cs.CL2026

Bridging the Editing Gap in LLMs: FineEdit for Precise and Targeted Text Modifications

Yiming Zeng, Wanhao Yu, Zexin Li +5

Large Language Models (LLMs) have significantly advanced natural language processing, demonstrating strong capabilities in tasks such as text generation, summarization, and reasoni…