3 papers
cs.AI2026
LongCat-Flash-Thinking-2601 Technical Report
Meituan LongCat Team, Anchun Gui, Bei Li +162
We introduce LongCat-Flash-Thinking-2601, a 560-billion-parameter open-source Mixture-of-Experts (MoE) reasoning model with superior agentic reasoning capability. LongCat-Flash-Thi…
hep-ex2025
New Physics Search at the CEPC: a General Perspective
Xiaocong Ai, Stefan Antusch, Peter Athron +211
The Circular Electron-Positron Collider (CEPC), a proposed next-generation Higgs factory, provides new opportunities to explore physics beyond the Standard Model (SM). With its cle…
cs.LG2024
Length Desensitization in Direct Preference Optimization
Wei Liu, Yang Bai, Chengcheng Han +5
Direct Preference Optimization (DPO) is widely utilized in the Reinforcement Learning from Human Feedback (RLHF) phase to align Large Language Models (LLMs) with human preferences,…