preprint Open access

Hilbert-Augmented Reinforcement Learning for Scalable Multi-Robot Coverage and Exploration

  • Open MIND
Research footprint

At a glance

Citations
0
References
0
Comments
0
Paper overview

Abstract

We present a coverage framework that integrates Hilbert space-filling priors into decentralized multi-robot learning and execution. We augment DQN and PPO with Hilbert-based spatial indices to structure exploration and reduce redundancy in sparse-reward environments, and we evaluate scalability in multi-robot grid coverage. We further describe a waypoint interface that converts Hilbert orderings into curvature-bounded, time-parameterized SE(2) trajectories (planar (x, y, θ)), enabling onboard feasibility on resource-constrained robots. Experiments show improvements in coverage efficiency, redundancy, and convergence speed over DQN/PPO baselines. In addition, we validate the approach on a Boston Dynamics Spot legged robot, executing the generated trajectories in indoor environments and observing reliable coverage with low redundancy. These results indicate that geometric priors improve autonomy and scalability for swarm and legged robotics.

Record transparency

Publication details

DOI
10.48550/arxiv.2602.19400
OpenAlex
W7131443712
Document type
preprint
Language
EN
Source
Open MIND
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.