2 papers
cs.LG2026
Escaping the Verifier: Learning to Reason via Demonstrations
Locke Cai, Max Ryabinin, Ivan Provilkov
Training Large Language Models (LLMs) to reason often relies on Reinforcement Learning (RL) with task-specific verifiers. However, many real-world reasoning-intensive tasks lack ve…
cs.LG2024
Identifying Money Laundering Subgraphs on the Blockchain
Kiwhan Song, Mohamed Ali Dhraief, Muhua Xu +4
Anti-Money Laundering (AML) involves the identification of money laundering crimes in financial activities, such as cryptocurrency transactions. Recent studies advanced AML through…