2 papers
cs.CL2026
Risk-Constrained Freshness-Aware Semantic Caching for Open-Web Retrieval-Augmented LLMs
Muhammad Mansoor, Tahir Ahmad, Yeo-Chan Yoon
Semantic caching reduces the latency and cost of retrieval-augmented generation (RAG) by serving cached answers to semantically similar queries, but most existing methods do not mo…
cs.CL2026
Sensory-Aware Sequential Recommendation via Review-Distilled Representations
Yeo Chan Yoon, Chanjun Park, Kyuhan Koh
We propose a novel framework for sensory-aware sequential recommendation that enriches item representations with linguistically extracted sensory attributes from product reviews. O…