3 papers
cs.CL2026
SenseJudge: Human-Centric Preference-Driven Judgment Framework
Rui Li, Junfeng Liu, Xiangwen Kong +2
Large Language Models (LLMs) as judges across various scenarios such as assessing model responses is becoming an increasingly accepted paradigm. However, existing judgment approach…
cs.AI2026
DELTAMEM: Incremental Experience Memory for LLM Agents via Residual Trees
Haoran Tan, Zeyu Zhang, Zhicheng Cao +2
Large Language Model (LLM)-based agents increasingly rely on memory to learn from experiences over continual interactions. However, storing experiences as independent, flat units l…
cs.CV2026
FineVision: Open Data Is All You Need
Luis Wiedmann, Orr Zohar, Amir Mahla +6
The advancement of vision-language models (VLMs) is hampered by a fragmented landscape of inconsistent and contaminated public datasets. We introduce FineVision, a meticulously col…