19 papers
TEngineDB-V: An OLAP-Native Vector Search System for Large- Workloads at Tencent
Xufei Wu, Pengcheng Zhang, Yitong Song +11
Vector search systems are essential infrastructure for modern data-driven applications. Large- analytical vector search, which retrieves -- results for analytics (…
Are We Ready For An Agent-Native Memory System?
Wei Zhou, Xuanhe Zhou, Shaokun Han +5
Memory for large language model (LLM) agents has rapidly evolved from simple retrieval-augmented mechanisms into a data management system that supports persistent information stora…
X+Slides: Benchmarking Audience-Conditioned Slide Generation
Haodong Chen, Xuanhe Zhou, Wei Zhou +6
Automatically generating slide decks from source documents is an important application of large language models (LLMs). Existing benchmarks primarily assess slide completeness and…
SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning
Chenzhi Hu, Qinzhe Hu, Yuhang Xu +6
Large reasoning models (LRMs) like OpenAI o1 and DeepSeek-R1 achieve high accuracy on complex tasks by adopting long chain-of-thought (CoT) reasoning paths. However, the inherent v…
SetupX: Can LLM Agents Learn from Past Failures in Functionality-Correct Code Repository Setup?
Zihang Zhou, Ziqian Ren, Yukai Wu +7
Functionality-correct repository setup aims to configure execution environments (e.g., dependencies, build scripts) to successfully execute a repository's documented features. It p…
An Efficient and Privacy-Preserving Architecture for Cross-Institutional Collaborative RAG
Chenxin Mao, Shangyu Liu, Zhenzhe Zheng +3
Retrieval-Augmented Generation (RAG) empowers LLMs with external knowledge, making cross-institutional domain-specific knowledge base integration a highly promising deployment para…