2 papers
cs.SE2026
A Scalable Benchmark for Repository-Oriented Long-Horizon Conversational Context Management
Yang Liu, Li Zhang, Fang Liu +2
In recent years, large language models (LLMs) have advanced rapidly, substantially enhancing their code understanding and generation capabilities and giving rise to powerful code a…
cs.IR2025
KAP: MLLM-assisted OCR Text Enhancement for Hybrid Retrieval in Chinese Non-Narrative Documents
Hsin-Ling Hsu, Ping-Sheng Lin, Jing-Di Lin +1
Hybrid Retrieval systems, combining Sparse and Dense Retrieval methods, struggle with Traditional Chinese non-narrative documents due to their complex formatting, rich vocabulary,…