3 papers
cs.LG2026
Granularity-Adaptive Credit Assignment for Long-Horizon LLM Agent Reinforcement Learning
Taoran Liang, Yang Liu, Shang Luo +9
Reinforcement learning is now the standard way to train large language model agents on long-horizon tasks, where dozens of interdependent actions precede a single sparse reward. Cr…
cs.SI2026
SAHG: Sector-Anisotropic Hyperbolic Graph Model for Social Bot Detection
Hanning Lu, Yingguang Yang, Jinwei Su +8
LLM-driven social bots can generate fluent, human-like text, reducing the discriminative advantage of content-based detection alone. However, coordinated campaigns still leave rela…
cs.AI2025
Introducing LongCat-Flash-Thinking: A Technical Report
Meituan LongCat Team, Anchun Gui, Bei Li +122
We present LongCat-Flash-Thinking, an efficient 560-billion-parameter open-source Mixture-of-Experts (MoE) reasoning model. Its advanced capabilities are cultivated through a metic…