papers

Publications (82)

math.OC2019

Optimal Resource Procurement and the Price of Causality

Sen Li, Akhil Shetty, Kameshwar Poolla +1

This paper studies the problem of procuring diverse resources in a forward market to cover a set of uncertain demand signals . We consider two scenarios: (a) $\bf{…

cs.IR2022

Multi-Objective Personalized Product Retrieval in Taobao Search

Yukun Zheng, Jiang Bian, Guanghao Meng +7

In large-scale e-commerce platforms like Taobao, it is a big challenge to retrieve products that satisfy users from billions of candidates. This has been a common concern of academ…

cs.CV2026

A welding penetration prediction model for laser welding process based on self-supervised learning using physics-informed neural networks

Sen Li, Xiaoying Liu, Xiaojian Xu +5

The laser welding full-penetration is of critical importance, as it constitutes one of the fundamental factors in achieving defect-free welded joints. Accurate prediction of the pe…

math.OC2024

Towards a Multimodal Charging Network: Joint Planning of Charging Stations and Battery Swapping Stations for Electrified Ride-Hailing Fleets

Zhijie Lai, Sen Li

This paper considers a multimodal charging network in which charging stations and battery swapping stations are jointly built to support an electric ride-hailing fleet synergistica…

econ.EM2021

Impact of Congestion Charge and Minimum Wage on TNCs: A Case Study for San Francisco

Sen Li, Kameshwar Poolla, Pravin Varaiya

This paper describes the impact on transportation network companies (TNCs) of the imposition of a congestion charge and a driver minimum wage. The impact is assessed using a market…

cs.MA2026

RideGym: A Standardized Interface for Real-World Large-Scale Ride-Sharing System

Zijian Zhao, Yulong Hu, Sen Li

Ride-sharing has become an essential component of modern urban transportation and has attracted significant attention across computer science, transportation, and management scienc…

eess.SP2026

Invertible Diffusion for Low-Memory Channel Gain Map Construction in Wireless Communication Networks

Ruifeng Gao, Sen Li, Jue Wang +2

Channel gain maps (CGMs) enable propagation-aware services in edge-intelligent wireless communication networks, while diffusion-based CGM construction is memory intensive for on-de…

cs.CV2026

A multi-task spatiotemporal deep neural network for predicting penetration depth and morphology in laser welding

Sen Li, Haichao Cui, Chendong Shao +2

In laser penetration welding, the assessment of penetration state and weld seam morphology plays a crucial role in determining the weld quality. This paper presents a comprehensive…

math.OC2021

On-Demand Valet Charging for Electric Vehicles: Economic Equilibrium, Infrastructure Planning and Regulatory Incentives

Zhijie Lai, Sen Li

Many city residents cannot install their private electric vehicle (EV) chargers due to the lack of dedicated parking spaces or insufficient grid capacity. This presents a significa…

cs.CV2023

Emotional Talking Head Generation based on Memory-Sharing and Attention-Augmented Networks

Jianrong Wang, Yaxin Zhao, Li Liu +3

Given an audio clip and a reference face image, the goal of the talking head generation is to generate a high-fidelity talking head video. Although some audio-driven methods of gen…

cs.CV2019

DADA-2000: Can Driving Accident be Predicted by Driver Attention? Analyzed by A Benchmark

Jianwu Fang, Dingxin Yan, Jiahuan Qiao +3

Driver attention prediction is currently becoming the focus in safe driving research community, such as the DR(eye)VE project and newly emerged Berkeley DeepDrive Attention (BDD-A)…

eess.SY2025

QSTAformer: A Quantum-Enhanced Transformer for Robust Short-Term Voltage Stability Assessment against Adversarial Attacks

Yang Li, Chong Ma, Yuanzheng Li +3

Short-term voltage stability assessment (STVSA) is critical for secure power system operation. While classical machine learning-based methods have demonstrated strong performance,…

math.OC2016

On Social Optima of Non-Cooperative Mean Field Games

Sen Li, Wei Zhang, Lin Zhao

This paper studies the connections between mean-field games and the social welfare optimization problems. We consider a mean field game in functional spaces with a large population…

quant-ph2025

Ultrahigh threshold nonstabilizer nonlinear quantum error correcting code

Maga Grafe, Kaixuan Zhou, Zaman Tekin +5

We introduce a novel type of quantum error correcting code, called the spinor code, based on spaces defined by total spin. The code is a nonstabilizer code, and is also a nonlinear…

cs.LG2021

Interpreting and Boosting Dropout from a Game-Theoretic View

Hao Zhang, Sen Li, Yinchao Ma +3

This paper aims to understand and improve the utility of the dropout operation from the perspective of game-theoretic interactions. We prove that dropout can suppress the strength…

math.OC2020

Off-Street Parking for TNC Vehicles to Reduce Cruising Traffic

Sen Li, Junjie Qin, Hai Yang +2

This paper considers off-street parking for the cruising vehicles of transportation network companies (TNCs) to reduce the traffic congestion. We propose a novel business that inte…

cs.CV2022

TJ4DRadSet: A 4D Radar Dataset for Autonomous Driving

Lianqing Zheng, Zhixiong Ma, Xichan Zhu +9

The next-generation high-resolution automotive radar (4D radar) can provide additional elevation measurement and denser point clouds, which has great potential for 3D sensing in au…

eess.SY2020

Selling Demand Response Using Options

Deepan Muthirayan, Dileep Kalathil, Sen Li +2

Wholesale electricity markets in many jurisdictions use a two-settlement structure: a day-ahead market for bulk power transactions and a real-time market for fine-grain supply-dema…

cs.RO2024

Task-Space Riccati Feedback based Whole Body Control for Underactuated Legged Locomotion

Shunpeng Yang, Zejun Hong, Sen Li +3

This manuscript primarily aims to enhance the performance of whole-body controllers(WBC) for underactuated legged locomotion. We introduce a systematic parameter design mechanism f…

cs.LG2021

Adversarial Learning for Incentive Optimization in Mobile Payment Marketing

Xuanying Chen, Zhining Liu, Li Yu +5

Many payment platforms hold large-scale marketing campaigns, which allocate incentives to encourage users to pay through their applications. To maximize the return on investment, i…

math.OC2023

Charging Autonomous Electric Vehicle Fleet for Mobility-on-Demand Services: Plug in or Swap out?

Jing Gao, Sen Li

This paper compares two prevalent charging strategies for electric vehicles, plug-in charging and battery swapping, to investigate which charging strategy is superior for electric…

cs.RO2023

Active Surface with Passive Omni-Directional Adaptation of Soft Polyhedral Fingers for In-Hand Manipulation

Sen Li, Fang Wan, Chaoyang Song

Track systems effectively distribute loads, augmenting traction and maneuverability on unstable terrains, leveraging their expansive contact areas. This tracked locomotion capabili…

cs.RO2026

Multi-Embodiment Robotic Retargeting via Guided Diffusion Model

Zhefeng Cao, Ben Liu, Shunpeng Yang +3

Motion retargeting for specific robot from existing motion datasets is one critical step in transferring motion patterns from human behaviors to and across various robots. However,…

cs.CR2023

DABS: Data-Agnostic Backdoor attack at the Server in Federated Learning

Wenqiang Sun, Sen Li, Yuchang Sun +1

Federated learning (FL) attempts to train a global model by aggregating local models from distributed devices under the coordination of a central server. However, the existence of…

math.OC2019

Transactive Energy System: Market-Based Coordination of Distributed Energy Resources

Sen Li, Jianming Lian, Antonio Conejo +1

Distributed energy resources (DER) provide significant value for renewable energy integration in modern power grids. However, unlocking this value requires complex design and coord…

quant-ph2026

Field-Trial Quantum Key Distribution with Qubit-Based Frame Synchronization

Rui Guan, Jingchun Yu, Zhaoyun Li +7

Quantum key distribution (QKD) is a cryptographic technique that uses quantum mechanical principles to enable secure key exchange. Practical deployment of QKD requires robust, cost…

math.OC2023

Regulating For-Hire Autonomous Vehicles for An Equitable Multimodal Transportation Network

Jing Gao, Sen Li

This paper assesses the equity impacts of for-hire autonomous vehicles (AVs) and investigates regulatory policies that promote spatial and social equity in future autonomous mobili…

cs.MA2026

BMG-Q: Localized Bipartite Match Graph Attention Q-Learning for Ride-Pooling Order Dispatch

Yulong Hu, Siyuan Feng, Sen Li

This paper introduces Localized Bipartite Match Graph Attention Q-Learning (BMG-Q), a novel Multi-Agent Reinforcement Learning (MARL) algorithm framework tailored for ride-pooling…

math.OC2018

Connections between Mean-Field Game and Social Welfare Optimization

Sen Li, Wei Zhang, Lin Zhao

This paper studies the connection between a class of mean-field games and a social welfare optimization problem. We consider a mean-field game in function spaces with a large popul…

cs.IR2025

C2T-ID: Converting Semantic Codebooks to Textual Document Identifiers for Generative Search

Yingchen Zhang, Ruqing Zhang, Jiafeng Guo +4

Designing document identifiers (docids) that carry rich semantic information while maintaining tractable search spaces is a important challenge in generative retrieval (GR). Popula…

cs.LG2025

Coordinating Ride-Pooling with Public Transit using Reward-Guided Conservative Q-Learning: An Offline Training and Online Fine-Tuning Reinforcement Learning Framework

Yulong Hu, Tingting Dong, Sen Li

This paper introduces a novel reinforcement learning (RL) framework, termed Reward-Guided Conservative Q-learning (RG-CQL), to enhance coordination between ride-pooling and public…

math.OC2016

Constrained Linear Quadratic Stackelberg Games with Applications in Demand Response

Sen Li, Wei Zhang, Jianming Lian +1

This paper studies a class of dynamic Stackelberg games under open-loop information structure with constrained linear agent dynamics and quadratic utility functions. We show two im…

cs.IR2025

Retrieval-in-the-Chain: Bootstrapping Large Language Models for Generative Retrieval

Yingchen Zhang, Ruqing Zhang, Jiafeng Guo +3

Generative retrieval (GR) is an emerging paradigm that leverages large language models (LLMs) to autoregressively generate document identifiers (docids) relevant to a given query.…

cs.CL2025

From General Reasoning to Domain Expertise: Uncovering the Limits of Generalization in Large Language Models

Dana Alsagheer, Yang Lu, Abdulrahman Kamal +7

Recent advancements in Large Language Models (LLMs) have demonstrated remarkable capabilities in various domains. However, effective decision-making relies heavily on strong reason…

eess.SY2024

Physical Informed-Inspired Deep Reinforcement Learning Based Bi-Level Programming for Microgrid Scheduling

Yang Li, Jiankai Gao, Yuanzheng Li +4

To coordinate the interests of operator and users in a microgrid under complex and changeable operating conditions, this paper proposes a microgrid scheduling model considering the…

cond-mat.str-el2019

Phenomenological Single-Particle Green's Function for the Pseudogap and Superconducting Phases of High-T Cuprates

Jian-Hao Zhang, Sen Li, Yao Ma +3

We present a phenomenological Green's function to characterize the superconducting and pseudogap phases of the cuprates based on a microscopic theory of doped Mott insulators. In t…

eess.SY2024

Enhancing Cyber-Resilience in Integrated Energy System Scheduling with Demand Response Using Deep Reinforcement Learning

Yang Li, Wenjie Ma, Yuanzheng Li +3

Optimally scheduling multi-energy flow is an effective method to utilize renewable energy sources (RES) and improve the stability and economy of integrated energy systems (IES). Ho…

cond-mat.supr-con2019

Anomalous doping evolution of nodal dispersion revealed by in-situ ARPES on continuously doped cuprates

Yigui Zhong, Jianyu Guan, Jin Zhao +9

We study the systematic doping evolution of nodal dispersions by in-situ angle-resolved photoemission spectroscopy on the continuously doped surface of a high-temperature supercond…

math.OC2021

Spatial Pricing in Ride-Sourcing Markets under a Congestion Charge

Sen Li, Hai Yang, Kameshwar Poolla +1

This paper studies the optimal spatial pricing for a ride-sourcing platform subject to a congestion charge. The platform determines the ride prices over the transportation network…

cs.CV2024

MuLan: Multimodal-LLM Agent for Progressive and Interactive Multi-Object Diffusion

Sen Li, Ruochen Wang, Cho-Jui Hsieh +2

Existing text-to-image models still struggle to generate images of multiple objects, especially in handling their spatial positions, relative sizes, overlapping, and attribute bind…

cs.LG2026

Learning Earthquake Wave Arrival Time Picking from Labels with Inaccuracies

Sen Li, Xu Yang, S. Mostafa Mousavi +5

Inaccurately labeled training data, or "label noise", poses a significant threat to the integrity of supervised machine learning models. This corruption directly degrades performan…

math.NA2024

A conditional normalizing flow for domain decomposed uncertainty quantification

Sen Li, Ke Li, Yu Liu +1

In this paper we present a conditional KRnet (cKRnet) based domain decomposed uncertainty quantification (CKR-DDUQ) approach to propagate uncertainties across different physical do…

physics.geo-ph2023

SeisT: A foundational deep learning model for earthquake monitoring tasks

Sen Li, Xu Yang, Anye Cao +4

Seismograms, the fundamental seismic records, have revolutionized earthquake research and monitoring. Recent advancements in deep learning have further enhanced seismic signal proc…

math.OC2019

Regulating TNCs: Should Uber and Lyft Set Their Own Rules?

Sen Li, Hamidreza Tavafoghi, Kameshwar Poolla +1

We evaluate the impact of three proposed regulations of transportation network companies (TNCs) like Uber, Lyft and Didi: (1) a minimum wage for drivers, (2) a cap on the number of…

cs.IR2023

MAKE: Vision-Language Pre-training based Product Retrieval in Taobao Search

Xiaoyang Zheng, Zilong Wang, Ke Xu +4

Taobao Search consists of two phases: the retrieval phase and the ranking phase. Given a user query, the retrieval phase returns a subset of candidate products for the following ra…

cs.CV2026

A cross-process welding penetration status prediction algorithm based on unsupervised domain adaptation in laser and TIG welding

Sen Li, Haichao Cui, Chendong Shao +2

Supervised deep learning has been widely used for weld penetration state classification; however, its performance often degrades significantly under domain shift, such as when tran…

cs.IR2023

Rethinking the Role of Pre-ranking in Large-scale E-Commerce Searching System

Zhixuan Zhang, Yuheng Huang, Dan Ou +4

E-commerce search systems such as Taobao Search, the largest e-commerce searching system in China, aim at providing users with the most preferred items (e.g., products). Due to the…

cs.CV2026

Car-1000: A New Large Scale Fine-Grained Visual Categorization Dataset

Yutao Hu, Sen Li, Jincheng Yan +2

Fine-grained visual categorization (FGVC) is a challenging but significant task in computer vision, which aims to recognize different sub-categories of birds, cars, airplanes, etc.…

cs.RO2022

Research on Event Accumulator Settings for Event-Based SLAM

Kun Xiao, Guohui Wang, Yi Chen +3

Event cameras are a new type of sensors that are different from traditional cameras. Each pixel is triggered asynchronously by event. The trigger event is the change of the brightn…

eess.SY2023

Piggyback on Idle Ride-Sourcing Drivers for Integrated On-Demand and Flexible Intracity Parcel Delivery Services

Yang Liu, Sen Li

This paper investigates the spatial pricing and fleet management strategies for an integrated platform that provides both ride-sourcing services and intracity parcel delivery servi…

cs.CV2026

Physics-Guided Spatiotemporal State Space Modeling for Lookahead Molten Pool Segmentation in Laser Wire-Feed Welding

Sen Li, Haichao Cui, Changhao Yin +4

Real-time weld-pool perception is critical for closed-loop control in laser wire-feed welding, where sensing, computation, and actuator response introduce unavoidable delay. This p…

cs.IR2025

LLMs as Sparse Retrievers:A Framework for First-Stage Product Search

Hongru Song, Yu-an Liu, Ruqing Zhang +6

Product search is a crucial component of modern e-commerce platforms, with billions of user queries every day. In product search systems, first-stage retrieval should achieve high…

math.OC2023

Spatiotemporal Pricing and Fleet Management of Autonomous Mobility-on-Demand Networks: A Decomposition and Dynamic Programming Approach with Bounded Optimality Gap

Zhijie Lai, Sen Li

This paper studies spatiotemporal pricing and fleet management for autonomous mobility-on-demand (AMoD) systems while taking elastic demand into account. We consider a platform tha…

math.NA2022

A deep domain decomposition method based on Fourier features

Sen Li, Yingzhi Xia, Yu Liu +1

In this paper we present a Fourier feature based deep domain decomposition method (F-D3M) for partial differential equations (PDEs). Currently, deep neural network based methods ar…

cs.SD2023

MAVD: The First Open Large-Scale Mandarin Audio-Visual Dataset with Depth Information

Jianrong Wang, Yuchen Huo, Li Liu +3

Audio-visual speech recognition (AVSR) gains increasing attention from researchers as an important part of human-computer interaction. However, the existing available Mandarin audi…

cond-mat.mes-hall2020

Spin orbit field in a physically defined p type MOS silicon double quantum dot

Marian Marx, Jun Yoneda, Ángel Gutiérrez Rubio +10

We experimentally and theoretically investigate the spin orbit (SO) field in a physically defined, p type metal oxide semiconductor double quantum dot in silicon. We measure the ma…

cs.LG2025

Triple-BERT: Do We Really Need MARL for Order Dispatch on Ride-Sharing Platforms?

Zijian Zhao, Sen Li

On-demand ride-sharing platforms, such as Uber and Lyft, face the intricate real-time challenge of bundling and matching passengers-each with distinct origins and destinations-to a…

cs.CV2024

Temporal Action Localization with Cross Layer Task Decoupling and Refinement

Qiang Li, Di Liu, Jun Kong +3

Temporal action localization (TAL) involves dual tasks to classify and localize actions within untrimmed videos. However, the two tasks often have conflicting requirements for feat…

cond-mat.mes-hall2018

Difference in charge and spin dynamics in a quantum dot-lead coupled system

Tomohiro Otsuka, Takashi Nakajima, Matthieu R. Delbecq +12

We analyze time evolution of charge and spin states in a quantum dot coupled to an electric reservoir. Utilizing high-speed single-electron detection, we focus on dynamics induced…

cs.CV2025

MetaOcc: Spatio-Temporal Fusion of Surround-View 4D Radar and Camera for 3D Occupancy Prediction with Dual Training Strategies

Long Yang, Lianqing Zheng, Wenjin Ai +8

Robust 3D occupancy prediction is essential for autonomous driving, particularly under adverse weather conditions where traditional vision-only systems struggle. While the fusion o…

quant-ph2026

Preparing squeezed, cat and GKP states with parity measurements

Zhiyuan Lin, Sen Li, Jingyan Feng +4

Bosonic modes constitute a central resource in a wide range of quantum technologies, providing long-lived degrees of freedom for the storage, processing, and transduction of quantu…

cs.MM2023

Memory-augmented Contrastive Learning for Talking Head Generation

Jianrong Wang, Yaxin Zhao, Li Liu +4

Given one reference facial image and a piece of speech as input, talking head generation aims to synthesize a realistic-looking talking head video. However, generating a lip-synchr…

cs.IR2023

Graph Contrastive Learning with Multi-Objective for Personalized Product Retrieval in Taobao Search

Longbin Li, Chao Zhang, Sen Li +3

In e-commerce search, personalized retrieval is a crucial technique for improving user shopping experience. Recent works in this domain have achieved significant improvements by th…

cs.LG2026

Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

Zijian Zhao, Jing Gao, Sen Li

Cooperative multi-agent reinforcement learning (MARL) is widely used to address large joint observation and action spaces by decomposing a centralized control problem into multiple…

cs.AI2025

CitySeeker: How Do VLMS Explore Embodied Urban Navigation With Implicit Human Needs?

Siqi Wang, Chao Liang, Yunfan Gao +5

Vision-Language Models (VLMs) have made significant progress in explicit instruction-based navigation; however, their ability to interpret implicit human needs (e.g., "I am thirsty…

cs.IR2022

Modeling User Behavior with Graph Convolution for Personalized Product Search

Fan Lu, Qimai Li, Bo Liu +7

User preference modeling is a vital yet challenging problem in personalized product search. In recent years, latent space based methods have achieved state-of-the-art performance b…

cond-mat.mes-hall2015

Formation of Long Single Quantum Dots in High Quality InSb Nanowires Grown by Molecular Beam Epitaxy

Dingxun Fan, Sen Li, N. Kang +6

We report on realization and transport spectroscopy study of single quantum dots (QDs) made from InSb nanowires grown by molecular beam epitaxy (MBE). The nanowires employed are 50…

cs.CR2024

A Watermark-Conditioned Diffusion Model for IP Protection

Rui Min, Sen Li, Hongyang Chen +1

The ethical need to protect AI-generated content has been a significant concern in recent years. While existing watermarking strategies have demonstrated success in detecting synth…

cs.MA2026

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization

Zijian Zhao, Sen Li

Multi-agent policy optimization, exemplified by PPO-based methods, is a key branch of cooperative Multi-Agent Reinforcement Learning (MARL). A central design question is how many n…

math.OC2016

Multi-Stage Pricing for Coordination of Thermostatically Controlled Loads: A Dynamic Stackelberg Game Approach

Sen Li, Wei Zhang, Jianming Lian +1

This paper focuses on multi-stage coordination for a population of thermostatically controlled loads (TCL). Each load maximizes the individual utility in response to an energy pric…

astro-ph.CO2025

Inference of -mode polarization in the presence of non-Gaussian foregrounds

Sen Li, Chang Feng, Filipe B. Abdalla

The inflationary -mode signals encode invaluable information about the origin of our Universe and searching for potential signatures of primordial gravitational waves (PGWs) is…

cs.LG2024

Invisible Backdoor Attacks on Diffusion Models

Sen Li, Junchi Ma, Minhao Cheng

In recent years, diffusion models have achieved remarkable success in the realm of high-quality image generation, garnering increased attention. This surge in interest is parallele…

cs.IR2021

Embedding-based Product Retrieval in Taobao Search

Sen Li, Fuyu Lv, Taiwei Jin +5

Nowadays, the product search service of e-commerce platforms has become a vital shopping channel in people's life. The retrieval phase of products determines the search system's qu…

math.OC2024

A Two-Stage Online Algorithm for EV Charging Station Energy Management and Carbon Trading

Dongxiang Yan, Shihan Huang, Sen Li +2

The increasing electric vehicle (EV) adoption challenges the energy management of charging stations (CSs) due to the large number of EVs and the underlying uncertainties. Moreover,…

eess.SY2025

Joint Infrastructure Planning and Order Assignment for On-Demand Food-Delivery Services with Coordinated Drones and Human Couriers

Yang Liu, Yitong Shang, Sen Li

This paper investigates the optimal infrastructure planning and order assignment problem of an on-demand food-delivery platform with a mixed fleet of drones and human couriers. The…

cs.LG2026

AutoFed: Personalized Federated Traffic Prediction via Adaptive Prompt

Zijian Zhao, Yitong Shang, Sen Li

Accurate traffic prediction is essential for Intelligent Transportation Systems, including ride-hailing, urban road planning, and vehicle fleet management. However, due to signific…

math.OC2023

Regulating Transportation Network Companies with a Mixture of Autonomous Vehicles and For-Hire Human Drivers

Di Ao, Jing Gao, Zhijie Lai +1

This paper investigates the equity impacts of autonomous vehicles (AV) on for-hire human drivers and passengers in a ride-hailing market, and examines regulation policies that prot…

eess.SY2015

Uniform-Price Mechanism Design for a Large Population of Dynamic Agents

Sen Li, Wei Zhang

This paper focuses on the coordination of a large population of dynamic agents with private information over multiple periods. Each agent maximizes the individual utility, while th…

physics.atom-ph2025

Pre-emptive parametric kill switch for evaporative atomic sources in vacuum

Shuang Li, Zhiyuan Lin, Sen Li +7

A robust pre-emptive kill switch for cold atom experiments is introduced to significantly reduce costly system reassembly or replacement. The design incorporates upper (alarm) and…

astro-ph.CO2026

Simultaneous inference of isotropic and anisotropic polarization rotation angles for next-generation cosmic microwave background experiments

Yue Zhang, Chang Feng, Sen Li +1

The ultrahigh-sensitivity polarization imaging of the cosmic microwave background (CMB) is a treasure trove for new physics. Searching for a predicted polarization angle rotation k…

math.OC2015

A Mechanism Design Approach for Coordination of Thermostatically Controlled Loads

Sen Li, Wei Zhang, Jianming Lian +1

This paper focuses on the coordination of a population of thermostatically controlled loads (TCLs) with unknown parameters to achieve group objectives. The problem involves designi…

cs.AI2026

One Step is Enough: Multi-Agent Reinforcement Learning based on One-Step Policy Optimization for Order Dispatch on Ride-Sharing Platforms

Zijian Zhao, Sen Li

Order dispatch is a critical task in ride-sharing systems with Autonomous Vehicles (AVs), directly influencing efficiency and profits. Recently, Multi-Agent Reinforcement Learning…