3 papers
cs.CL2026
Total Recall at What Cost? Benchmarking the Serving Cost of Agentic Memory Systems
Natchanon Pollertlam, Witchayut Kornsuwannawit
Long-running conversational agents increasingly rely on a memory system to avoid resending the whole conversation each turn, yet how much that costs to serve has received little sy…
cs.CL2026
Beyond the Context Window: A Cost-Performance Analysis of Fact-Based Memory vs. Long-Context LLMs for Persistent Agents
Natchanon Pollertlam, Witchayut Kornsuwannawit
Persistent conversational AI systems face a choice between passing full conversation histories to a long-context large language model (LLM) and maintaining a dedicated memory syste…
cs.CL2025
JAI-1: A Thai-Centric Large Language Model
Attapol T. Rutherford, Jullajak Karnjanaekarin, Narongkorn Panitsrisit +3
This technical report introduces JAI-1, a Thai-centric language model with 75B parameters. Recent Thai models have primarily relied on existing open-source models, applying additio…