agentic reinforcement learning 1hierarchical reward modeling 1sandbox simulation 1tool-use agents 1travel planning 1
From the 1 of 12 linked papers with an AI index.
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
WavBench: Benchmarking Reasoning, Colloquialism, and Paralinguistics for End-to-End Spoken Dialogue Models
Yangzhuo Li, Shengpeng Ji, Yifu Chen +6
With the rapid integration of advanced reasoning capabilities into spoken dialogue models, the field urgently demands benchmarks that transcend simple interactions to address real-…
cs.CL2025
DiMA: An LLM-Powered Ride-Hailing Assistant at DiDi
Yansong Ning, Shuowei Cai, Wei Li +4
On-demand ride-hailing services like DiDi, Uber, and Lyft have transformed urban transportation, offering unmatched convenience and flexibility. In this paper, we introduce DiMA, a…
cs.CL2025
Not All Thoughts are Generated Equal: Efficient LLM Reasoning via Multi-Turn Reinforcement Learning
Yansong Ning, Wei Li, Jun Fang +2
Compressing long chain-of-thought (CoT) from large language models (LLMs) is an emerging strategy to improve the reasoning efficiency of LLMs. Despite its promising benefits, exist…