Back to Research papers
Research paper index

Learning to Race in Minutes: Infoprop Dyna on the Mini Wheelbot

Devdutt Subhasish, Henrik Hose, Sebastian Trimpe

arXiv:2605.01096Published May 1, 20260 citations
  • cs.LG
  • cs.RO
  • reinforcement learning
  • sim-to-real
  • robot
  • action

Abstract

Reinforcement Learning (RL) has the potential to enable robots with fast, nonlinear, and unstable dynamics to reach the limits of their performance. However, most recent advances rely on carefully designed physics-based simulators and domain randomization to achieve successful sim-to-real transfer within reasonable wall-clock time. In this work, we bypass the need for such simulators and demonstrate that Infoprop Dyna, a state-of-the-art uncertainty-aware model-based reinforcement learning (MBRL) framework, can enable robots to learn directly from real-world interactions. Using Infoprop Dyna, the Mini Wheelbot, an underactuated unicycle robot, learns to race around a track within 11 minutes of real-world experience.

Read the original paper

This page indexes public paper metadata. The manuscript remains with its original publisher and authors.