3 papers
cs.NI2026
Supercharging Packet-level Network Simulation of Large Model Training via Memoization and Fast-Forwarding
Fei Long, Kaihui Gao, Li Chen +6
Packet-level discrete-event simulation (PLDES) is a prevalent tool for evaluating detailed performance of large model training. Although PLDES offers high fidelity and generality,…
cs.LG2026
TTCS: Test-Time Curriculum Synthesis for Self-Evolving
Chengyi Yang, Zhishang Xiang, Yunbo Tang +5
Test-Time Training offers a promising way to improve the reasoning ability of large language models (LLMs) by adapting the model using only the test questions. However, existing me…
cs.AI2025
DMA: Online RAG Alignment with Human Feedback
Yu Bai, Yukai Miao, Dawei Wang +9
Retrieval-augmented generation (RAG) systems often rely on static retrieval, limiting adaptation to evolving intent and content drift. We introduce Dynamic Memory Alignment (DMA),…