activity
20242026
collaborators

5 papers

cs.LG2026

Monte Carlo Tree Search for Execution-Guided Program Repair with Large Language Models

Yixuan Liang

Automated program repair with large language models remains challenging at the repository level due to long-horizon reasoning requirements and the limitations of autoregressive dec…

cs.AI2025

Budget-Aware Tool Use Enables Effective Agent Scaling

Tengxiao Liu, Zifeng Wang, Jin Miao +12

Scaling test-time computation has been extended from language model reasoning to tool-augmented agents, where scaling involves not only thinking in tokens but also acting via tool…

cs.LG2025

Outlier Weighed Layerwise Sparsity (OWL): A Missing Secret Sauce for Pruning LLMs to High Sparsity

Lu Yin, You Wu, Zhenyu Zhang +10

Large Language Models (LLMs), renowned for their remarkable performance across diverse domains, present a challenge when it comes to practical deployment due to their colossal mode…

cs.CL2025

Boosting Reward Model with Preference-Conditional Multi-Aspect Synthetic Data Generation

Jiaming Shen, Ran Xu, Yennie Jun +6

Reward models (RMs) are crucial for aligning large language models (LLMs) with human preferences. They are trained using preference datasets where each example consists of one inpu…

cs.CL2024

Integrating Planning into Single-Turn Long-Form Text Generation

Yi Liang, You Wu, Honglei Zhuang +8

Generating high-quality, in-depth textual documents, such as academic papers, news articles, Wikipedia entries, and books, remains a significant challenge for Large Language Models…