2 papers
cs.CL2024
Toward Optimal Search and Retrieval for RAG
Alexandria Leto, Cecilia Aguerrebere, Ishwar Bhati +3
Retrieval-augmented generation (RAG) is a promising method for addressing some of the memory-related challenges associated with Large Language Models (LLMs). Two separate systems f…
cs.IR2024
GleanVec: Accelerating vector search with minimalist nonlinear dimensionality reduction
Mariano Tepper, Ishwar Singh Bhati, Cecilia Aguerrebere +1
Embedding models can generate high-dimensional vectors whose similarity reflects semantic affinities. Thus, accurately and timely retrieving those vectors in a large collection tha…