14 citations · 17 across the 13 of their papers we have counts for
15 papers
A Tri-Agent Framework for Evaluating and Aligning Question Clarification Capabilities of Large Language Models
Yikai Zhao, Saurabh Pandey, Pradeep Kumar Misra
Large Language Models (LLMs) are increasingly deployed in interactive systems where understanding user intent precisely is paramount. A key capability for such systems is effective…
Context-Aware Cluster Decoding: Semantic Anchor-Driven Coherence in dMLLMs
Yikai Zhao, Qiyan Zhao, Jiaquan Zhang +3
Diffusion multimodal large language models (dMLLMs) frequently produce long-form outputs marred by semantic drift and repetition, with quality generally degrading as output length…
TRACE: TRajectory Attribution for Automated Context Engineering
Yikai Zhao, Pradeep Kumar Misra, Saurabh Pandey
Production AI agents fail when their context sources -- system prompts, knowledge bases, tool descriptions, and procedural skills -- contain errors or gaps. Current maintenance rel…
Kimi K3: Open Frontier Intelligence
Kimi Team, Tongtong Bai, Yifan Bai +398
We introduce Kimi K3, a 2.8T parameter Mixture-of-Experts model with 104 billion activated parameters, native vision capabilities, and a 1-million-token context window. Kimi K3 is…
TENT: A Declarative Slice Spraying Engine for Performant and Resilient Data Movement in Disaggregated LLM Serving
Feng Ren, Ruoyu Qin, Teng Ma +16
Modern GPU clusters rely on complex, heterogeneous interconnects. As large language model (LLM) serving shifts toward agentic reasoning, KVCache becomes a first-class mobile asset,…
Kimi K2.5: Visual Agentic Intelligence
Kimi Team, Tongtong Bai, Yifan Bai +333
We introduce Kimi K2.5, an open-source multimodal agentic model designed to advance general agentic intelligence. K2.5 emphasizes the joint optimization of text and vision so that…