6 citations · 9 across the 7 of their papers we have counts for
3 papers · 1 filter
Graph-based Target Back-Propagation for Context Adaptation in Multi-LLM Agentic Systems
Tan Zhu, Tong Yao, Kananart Kuwaranancharoen +4
Context adaptation automates prompt engineering in LLM-based systems by iteratively revising tunable prompts from task feedback, without modifying model weights. Extending this par…
The Serial Scaling Hypothesis
Yuxi Liu, Konpat Preechakul, Kananart Kuwaranancharoen +1
While machine learning has advanced through massive parallelization, we identify a critical blind spot: some problems are fundamentally sequential. These "inherently serial" proble…
On the Benefits of Leveraging Structural Information in Planning Over the Learned Model
Jiajun Shen, Kananart Kuwaranancharoen, Raid Ayoub +2
Model-based Reinforcement Learning (RL) integrates learning and planning and has received increasing attention in recent years. However, learning the model can incur a significant…