works on

From the 1 of 41 linked papers with an AI index.

activity
20242026
collaborators

41 papers

cs.AI2026

Agent-MD: Selective LLM Intervention with Event-Driven Escalation for Stateful GCMC--MD Campaigns

Yijie Wang, Zhen-Yu Yin, Zhenheng Tang +1

Long-running molecular simulation campaigns require repeated continuation from saved states, provenance-aware progression, adaptive assessment, and occasional interpretation of wor…

cs.LG2026

SNI-GNN: SmartNIC-Assisted Full-Graph GNN Training with In-Network Embedding Prediction

Guofan Yu, Sitian Chen, Zhenheng Tang +2

Full-graph GNN training delivers high accuracy but scales poorly on multi-server clusters due to heavy, irregular inter-node embedding exchanges. We present SNI-GNN, a SmartNIC-ass…

cs.DC2026

Zellige: Moldable Sequence Placement for Mixed Image-Video DiT Training

Guangyu Xiang, Xueze Kang, Minwei Zhao +4

High-quality video generation requires training Diffusion Transformers (DiTs) jointly on image and video data, posing a mixed-length sequence training problem across GPUs. Existing…

cs.LG2026

DRNOISE: Benchmarking Deep Research Agents in Misleading Evidence Environments

Jun Nie, Zhiqin Yang, Zhenheng Tang +4

Deep research agents increasingly operate over the open web, where relevant records coexist with redundant summaries, outdated reports, and misleading documents. Existing evaluatio…

cs.DC2026

Xema: Efficient Diffusion Serving through Fine-Grained Memory Management and Auto-Configuration

Xueze Kang, Guangyu Xiang, Suyi Li +4

Xema is a system that reduces GPU memory usage for diffusion model serving by analyzing tensor lifetimes to apply targeted memory mitigation and by planning parallelism and concurr…

cs.DC2026

Bidirectional Resource Scheduling for Disaggregated and Asynchronous RL Post-Training

Tan Zhiqiang, Zhiqiang Tan, Maoxin Wang +11

It is well established that the reasoning capabilities of large language models (LLMs) can be improved by applying reinforcement learning (RL) in a post-training stage. In a standa…