Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Playing Psychic: Using Thought Trees to Predict Reasoning Models Accuracy on Coding Tasks
Jiaxin Fang, Runyuan He, Sahil Bhatia +2
Recent advances in large language models (LLMs) have shown that test-time scaling can substantially improve model performance on complex tasks, particularly in the coding domain. U…
cs.AI2026
Combee: Scaling Prompt Learning for Self-Improving Language Model Agents
Hanchen Li, Runyuan He, Qizheng Zhang +11
Recent advances in prompt learning allow large language model agents to acquire task-relevant knowledge from inference-time context without parameter changes. For example, existing…