2 papers
cs.AI2026
DIG to Heal: Scaling General-purpose Agent Collaboration via Explainable Dynamic Decision Paths
Hanqing Yang, Hyungwoo Lee, Yuhang Yao +4
The increasingly popular agentic AI paradigm promises to harness the power of multiple, general-purpose large language model (LLM) agents to collaboratively complete complex tasks.…
cs.AI2025
Shared Disk KV Cache Management for Efficient Multi-Instance Inference in RAG-Powered LLMs
Hyungwoo Lee, Kihyun Kim, Jinwoo Kim +5
Recent large language models (LLMs) face increasing inference latency as input context length and model size continue to grow. In particular, the retrieval-augmented generation (RA…