3 papers
cs.AI2026
RLCascadeRouter: Quality-Estimator-Free Cascade Routing via Reinforcement Learning
Shihong Huang, Shengjie Wang, Hong Ma +1
The growing ecosystem of large language models (LLMs) offers huge potential to optimize performance-cost trade-offs. However, their heterogeneous capabilities and inference costs m…
cs.LG2026
Vehicle-as-Prompt: A Unified Deep Reinforcement Learning Framework for Heterogeneous Fleet Vehicle Routing Problem
Shihong Huang, Shengjie Wang, Lei Gao +4
Unlike traditional homogeneous routing problems, the Heterogeneous Fleet Vehicle Routing Problem (HFVRP) involves heterogeneous fixed costs, variable travel costs, and capacity con…
cs.DS2025
Potential-Based Greedy Matching for Dynamic Delivery Pooling
Hongyao Ma, Will Ma, Matias Romero
We study the dynamic pooling of multiple orders into a single trip, a strategy widely adopted by online delivery platforms. When an order has to be dispatched, the platform must de…