collaborators

15 papers

cs.AI2026

CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation

Yaoning Yu, Kai-Min Chang, Ye Yu +3

Online credit card discussions provide a natural setting for studying how consumers communicate about financial products. Simulating these discussions requires more than just gener…

cs.AI2026

KV-PRM: Efficient Process Reward Modeling via KV-Cache Transfer for Multi-Agent Test-Time Scaling

Peng Kuang, Haibo Jin, Xiaoyu Han +5

Process Reward Models (PRMs) have been proven to be highly effective in guiding test-time scaling (TTS) methods, which significantly boost the capabilities of LLM-based multi-agent…

cs.CL2026

Nemotron-Labs-Diffusion: A Tri-Mode Language Model Unifying Autoregressive, Diffusion, and Self-Speculation Decoding

Yonggan Fu, Lexington Whalen, Abhinav Garg +23

We introduce Nemotron-Labs-Diffusion, a tri-mode language model (LM) that unifies AR, diffusion, and self-speculation decoding within a single architecture. Trained with a joint AR…

cs.AI2026

Closing the Loop on Latent Reasoning via Test-Time Reconstruction

Xiaopeng Yuan, Haibo Jin, Ye Yu +4

Recent work moves intermediate reasoning from natural-language traces into latent or cache-level representations to reduce token overhead and avoid a discrete communication bottlen…

cs.MA2026

Agent Primitives: Reusable Latent Building Blocks for Multi-Agent Systems

Haibo Jin, Peng Kuang, Ye Yu +2

While existing multi-agent systems (MAS) can handle complex problems by enabling collaboration among multiple agents, they are often highly task-specific, relying on manually craft…

cs.MA2026

MiroBench: Benchmarking Realism in Agentic Simulation of Real-world Discussions

Yaoning Yu, Ye Yu, Haojing Luo +1

LLM agents are increasingly used to simulate real world interactions, but it remains unclear whether simulated behaviors preserve the content patterns and interaction dynamics of r…