2 papers
cs.CL2026
LoopRPT: Reinforcement Pre-Training for Looped Language Models
Guo Tang, Shixin Jiang, Heng Chang +6
Looped language models (LoopLMs) perform iterative latent computation to refine internal representations, offering a promising alternative to explicit chain-of-thought (CoT) reason…
cs.CL2025
Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering
Zheng Chu, Huiming Fan, Jingchang Chen +8
Although large language models (LLMs) have demonstrated remarkable reasoning capabilities, they still face challenges in knowledge-intensive multi-hop reasoning. Recent work explor…