most citedMiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling

1 citations · 1 across the 16 of their papers we have counts for

collaborators
Showing cs.AIShow all

5 papers · 1 filter

cs.AI2026

HippoSpark: An On-Demand Experience System for LLM Reasoning

Jingyao Liu, Danling Meng, Chen Huang +5

Distilling historical trajectories into reusable experience to enhance future problem-solving has become a focal point of recent LLM research. However, existing methods predominant…

cs.AI2026

AgentCPM-Report: Interleaving Drafting and Deepening for Open-Ended Deep Research

Yishan Li, Wentong Chen, Yukun Yan +12

Generating deep research reports requires large-scale information acquisition and the synthesis of insight-driven analysis, posing a significant challenge for current language mode…

cs.AI2026

AgentCPM-Explore: Realizing Long-Horizon Deep Exploration for Edge-Scale Agents

Haotian Chen, Xin Cong, Shengda Fan +16

While Large Language Model (LLM)-based agents have shown remarkable potential for solving complex tasks, existing systems remain heavily reliant on large-scale models, leaving the…

cs.AI2026

Empirical Analysis of Decoding Biases in Masked Diffusion Models

Pengcheng Huang, Tianming Liu, Zhenghao Liu +5

Masked diffusion models (MDMs), which leverage bidirectional attention and a denoising process, are narrowing the performance gap with autoregressive models (ARMs). However, their…

cs.AI2025

Benchmarking Retrieval-Augmented Generation in Multi-Modal Contexts

Zhenghao Liu, Xingsheng Zhu, Tianshuo Zhou +5

With the rapid advancement of Multi-modal Large Language Models (MLLMs), their capability in understanding both images and text has greatly improved. However, their potential for l…