2 papers
cs.CL2026
Scaling LLM Knowledge Boundaries via Distribution-Optimized Synthesis
Songze Li, Yarong Lan, Zhongpu Bo +16
Knowledge injection via synthetic data is crucial for enhancing Large Language Models (LLMs). However, current synthesis methods simply stop at preset token counts or fixed data ra…
cs.AI2026
STAR: SpatioTemporal Adaptive Reward Allocation for Text-to-Image RL Post-Training
Jinjie Shen, Wei Deng, Xian Hu +2
Existing RL post-training methods for text-to-image generation usually convert the final-image reward into a single scalar advantage and apply it with the same strength to the enti…