From the 1 of 4 linked papers with an AI index.
4 papers
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval
Suhyeong Park, Junha Jung, Jungwoo Park +1
The paper introduces SaMer, an object‑aware token merging method that compresses image tokens into a small set of centroids while keeping the late‑interaction interface for vision‑…
From Voting to Agent Collaboration: Answer-Type-Aware LLM Pipelines for BioASQ 14b
Taeyun Roh, Eunha Lee, Wonjune Jang +3
Biomedical question answering requires not only accurate extraction of information from scientific literature but also reliable integration of evidence across multiple documents. T…
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning
Junha Jung, Minbyul Jeong, Suhyeon Lim +5
Recent multimodal large language models have shown great promise in clinical image reasoning, but existing post-training pipelines remain predominantly outcome-centric, relying on…
CLAG: Adaptive Memory Organization via Agent-Driven Clustering for Small Language Model Agents
Taeyun Roh, Wonjune Jang, Junha Jung +1
Large language model agents heavily rely on external memory to support knowledge reuse and complex reasoning tasks. Yet most memory systems store experiences in a single global ret…