activity
20232026
most citedStarling: An I/O-Efficient Disk-Resident Graph Index Framework for High-Dimensional Vector Similarity Search on Data Segment

62 citations · 93 across the 19 of their papers we have counts for

collaborators
Showing 2025 · cs.DBShow all

8 papers · 2 filters

cs.DB2025

Beyond Relational: Semantic-Aware Multi-Modal Analytics with LLM-Native Query Optimization

Junhao Zhu, Lu Chen, Xiangyu Ke +4

Multi-modal analytical processing has the potential to transform applications in e-commerce, healthcare, entertainment, and beyond. However, real-world adoption remains elusive due…

cs.DB2025

Scalable Graph Indexing using GPUs for Approximate Nearest Neighbor Search

Zhonggen Li, Xiangyu Ke, Yifan Zhu +3

Approximate nearest neighbor search (ANNS) in high-dimensional vector spaces has a wide range of real-world applications. Numerous methods have been proposed to handle ANNS efficie…

cs.DB2025

Balancing the Blend: An Experimental Analysis of Trade-offs in Hybrid Search

Mengzhao Wang, Boyu Tan, Yunjun Gao +5

Hybrid search, the integration of lexical and semantic retrieval, has become a cornerstone of modern information retrieval systems, driven by demanding applications like Retrieval-…

cs.DB2025

OneDB: A Distributed Multi-Metric Data Similarity Search System

Tang Qian, Yifan Zhu, Lu Chen +5

Increasingly massive volumes of multi-modal data are being accumulated in many {real world} settings, including in health care and e-commerce. This development calls for effective…

cs.DB2025

Empowering Graph-based Approximate Nearest Neighbor Search with Adaptive Awareness Capabilities

Jiancheng Ruan, Tingyang Chen, Renchi Yang +2

Approximate Nearest Neighbor Search (ANNS) in high-dimensional spaces finds extensive applications in databases, information retrieval, recommender systems, etc. While graph-based…

cs.DB2025

Stitching Inner Product and Euclidean Metrics for Topology-aware Maximum Inner Product Search

Tingyang Chen, Cong Fu, Xiangyu Ke +3

Maximum Inner Product Search (MIPS) is a fundamental challenge in machine learning and information retrieval, particularly in high-dimensional data applications. Existing approache…