2 papers
cs.AI2026
KV-PRM: Efficient Process Reward Modeling via KV-Cache Transfer for Multi-Agent Test-Time Scaling
Peng Kuang, Haibo Jin, Xiaoyu Han +5
Process Reward Models (PRMs) have been proven to be highly effective in guiding test-time scaling (TTS) methods, which significantly boost the capabilities of LLM-based multi-agent…
cs.AI2026
Hawk: Harnessing Hardware-Aware Knowledge for High-Performance NPU Kernel Generation
Junyi Wen, Ruiyan Zhuang, Yongjia Xu +7
Developing high-performance kernels for Neural Processing Units (NPUs) is a critical industry bottleneck, requiring developers to manually navigate implicit hardware constraints an…