Model-Based Reinforcement Learning for Heterogeneous Multi-Robot Task Assignment Under Distribution Shifts

📅 2026-08-21
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
研究针对异构多机器人任务分配中分布变化的问题,提出了一种基于模型的强化学习方法,结合预测意识的自适应策略来提高服务效率和鲁棒性。
📝 Abstract
Heterogeneous multi-robot service systems must assign requests to compatible robots, construct feasible schedules, and adapt as new tasks arrive online. Historical data can help anticipate future demand, but relying too heavily on inaccurate predictions can degrade performance under distribution shifts. We develop a prediction-aware adaptive rollout framework for heterogeneous multi-robot task assignment with scheduled and real-time requests. The problem is formulated as a finite-horizon stochastic dynamic program incorporating robot-task compatibility, ordered service requirements, routing constraints, service windows, and end-of-horizon return requirements. The proposed policy evaluates current assignments using sampled future request scenarios while restricting immediate commitments to requests already observed. To enable online use, the framework combines pruned candidate controls, wait actions, and an interaction-aware base policy for efficient future-cost estimation. Robustness to forecast error is provided by adaptively reweighting predicted requests based on recent prediction mismatch and selectively re-optimizing assigned but unstarted requests. We also introduce a historical-data-driven procedure for selecting the heterogeneous fleet composition before deployment. In a case study using real nursing-task requests from hospital inpatient floors, the proposed approach achieves near-complete service and reduces serviced-request wait times relative to reactive, token-passing, prediction-positioning, and myopic greedy baselines, with the largest improvements in tail-delay metrics.
Problem

Research questions and friction points this paper is trying to address.

Heterogeneous Multi-Robot Systems
Task Assignment
Distribution Shifts
Adaptive Scheduling
Forecast Error
Innovation

Methods, ideas, or system contributions that make the work stand out.

Model-Based Reinforcement Learning
Heterogeneous Multi-Robot Systems
Prediction-Aware Adaptive Rollout
Robustness to Forecast Error
Dynamic Re-Optimization
💼 Related Jobs
No related jobs found.
D
Daniel Garces
S
Sara Castro
A
Adrian Haimovich
B
Byron Crowe
Stephanie Gil
Stephanie Gil
Assistant Professor, Harvard University
Networked roboticsmulti-robot control