ResearchTrend.AI
  • Papers
  • Communities
  • Events
  • Blog
  • Pricing
Papers
Communities
Social Events
Terms and Conditions
Pricing
Parameter LabParameter LabTwitterGitHubLinkedInBlueskyYoutube

© 2025 ResearchTrend.AI, All rights reserved.

  1. Home
  2. Papers
  3. 2011.03085
14
5

RealAnt: An Open-Source Low-Cost Quadruped for Education and Research in Real-World Reinforcement Learning

5 November 2020
Rinu Boney
Jussi Sainio
M. Kaivola
Arno Solin
Arno Solin
ArXivPDFHTML
Abstract

Current robot platforms available for research are either very expensive or unable to handle the abuse of exploratory controls in reinforcement learning. We develop RealAnt, a minimal low-cost physical version of the popular `Ant' benchmark used in reinforcement learning. RealAnt costs only ∼\sim∼350 EUR (\410)inmaterialsandcanbeassembledinlessthananhour.Wevalidatetheplatformwithreinforcementlearningexperimentsandprovidebaselineresultsonasetofbenchmarktasks.WedemonstratethattheRealAntrobotcanlearntowalkfromscratchfromlessthan10minutesofexperience.Wealsoprovidesimulatorversionsoftherobot(withthesamedimensions,state−actionspaces,anddelayednoisyobservations)intheMuJoCoandPyBulletsimulators.Weopen−sourcehardwaredesigns,supportingsoftware,andbaselineresultsforeducationaluseandreproducibleresearch.410) in materials and can be assembled in less than an hour. We validate the platform with reinforcement learning experiments and provide baseline results on a set of benchmark tasks. We demonstrate that the RealAnt robot can learn to walk from scratch from less than 10 minutes of experience. We also provide simulator versions of the robot (with the same dimensions, state-action spaces, and delayed noisy observations) in the MuJoCo and PyBullet simulators. We open-source hardware designs, supporting software, and baseline results for educational use and reproducible research.410)inmaterialsandcanbeassembledinlessthananhour.Wevalidatetheplatformwithreinforcementlearningexperimentsandprovidebaselineresultsonasetofbenchmarktasks.WedemonstratethattheRealAntrobotcanlearntowalkfromscratchfromlessthan10minutesofexperience.Wealsoprovidesimulatorversionsoftherobot(withthesamedimensions,state−actionspaces,anddelayednoisyobservations)intheMuJoCoandPyBulletsimulators.Weopen−sourcehardwaredesigns,supportingsoftware,andbaselineresultsforeducationaluseandreproducibleresearch.

View on arXiv
Comments on this paper