collaborators

19 papers

cs.DB2026

TEngineDB-V: An OLAP-Native Vector Search System for Large- Workloads at Tencent

Xufei Wu, Pengcheng Zhang, Yitong Song +11

Vector search systems are essential infrastructure for modern data-driven applications. Large- analytical vector search, which retrieves -- results for analytics (…

cs.CL2026

Are We Ready For An Agent-Native Memory System?

Wei Zhou, Xuanhe Zhou, Shaokun Han +5

Memory for large language model (LLM) agents has rapidly evolved from simple retrieval-augmented mechanisms into a data management system that supports persistent information stora…

cs.AI2026

X+Slides: Benchmarking Audience-Conditioned Slide Generation

Haodong Chen, Xuanhe Zhou, Wei Zhou +6

Automatically generating slide decks from source documents is an important application of large language models (LLMs). Existing benchmarks primarily assess slide completeness and…

cs.CL2026

SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning

Chenzhi Hu, Qinzhe Hu, Yuhang Xu +6

Large reasoning models (LRMs) like OpenAI o1 and DeepSeek-R1 achieve high accuracy on complex tasks by adopting long chain-of-thought (CoT) reasoning paths. However, the inherent v…

cs.SE2026

SetupX: Can LLM Agents Learn from Past Failures in Functionality-Correct Code Repository Setup?

Zihang Zhou, Ziqian Ren, Yukai Wu +7

Functionality-correct repository setup aims to configure execution environments (e.g., dependencies, build scripts) to successfully execute a repository's documented features. It p…

cs.CR2026

An Efficient and Privacy-Preserving Architecture for Cross-Institutional Collaborative RAG

Chenxin Mao, Shangyu Liu, Zhenzhe Zheng +3

Retrieval-Augmented Generation (RAG) empowers LLMs with external knowledge, making cross-institutional domain-specific knowledge base integration a highly promising deployment para…