microduck_rl

RL training environments for Microduck (mjlab)

72 articles 1.1k View on GitHub ↗
72 articles
Purpose of the `_servo_joint_ids()` Helper in Microduck RL: How It Isolates Active Joints in Isaac Gym Simulations

Discover how Microduck RL's _servo_joint_ids() helper isolates active joints, preventing accidental manipulation of physics elements and ensuring accurate reward functions in Isaac Gym simulations.

internals
Sep 8, 2026
How Backlash Is Modeled in the Microduck Robot for Sim2real Transfer

Learn how Microduck models backlash for sim2real transfer using hinge joints, a custom encoder actuator, and environment wrappers. Improve robot control and performance.

how-to-guide
Sep 8, 2026
How Microduck RL Handles NaN Values to Prevent Training Crashes

Learn how Microduck RL prevents training crashes from NaN values using four mechanisms: reward sanitisation, advantage sanitisation, observation NaN policies, and nan-state termination.

how-to-guide
Sep 8, 2026
Microduck RL MJCF Robot Models: Complete Guide to All 8 XML Configurations

Explore the 8 Microduck RL MJCF robot models for walking, ground-contact, roller-skate, and backlash variants. Access all configurations in this comprehensive guide.

getting-started
Sep 8, 2026
How to Publish Microduck RL Policies to HuggingFace Hub

Easily publish Microduck RL policies to HuggingFace Hub. Export ONNX, generate manifests, and upload with a single command using uv run publish.

how-to-guide
Sep 8, 2026
How to Export Microduck RL Checkpoints: A Complete Guide to ONNX Conversion

Export Microduck RL checkpoints easily with the official CLI script. Convert trained models to ONNX format, embedding the observation normalizer for seamless deployment. Learn how now.

how-to-guide
Sep 8, 2026
How the Observation Normalizer Is Baked Into the ONNX Export for Microduck RL

Learn how Microduck RL bakes the observation normalizer into ONNX export using the EmpiricalNormalization layer for seamless model integration and improved performance.

internals
Sep 8, 2026
Microduck RL Domain Randomization Strategy: How Non‑Accumulating Perturbations Enable Sim‑to‑Real Transfer

Discover the Microduck RL domain randomization strategy. Learn how non-accumulating perturbations ensure effective sim-to-real transfer for robotics.

deep-dive
Sep 8, 2026
How to Train a Microduck RL Policy with Backlash Simulation: Complete Guide

Learn to train a Microduck RL policy with backlash simulation. Follow our guide to select tasks, run training, and export your policy for hardware deployment. Get started today!

how-to-guide
Sep 8, 2026
Episodic Trick Tasks in Microduck RL: Forward Roll, Ground Pick, and Ball Kick Explained

Explore Microduck RL's episodic trick tasks like Forward Roll, Ground Pick, and Ball Kick. Learn about these fixed-duration maneuvers for robot learning.

tutorial
Sep 8, 2026
Microduck RL Locomotion Tasks: Complete Task Catalog for Bipedal Robot Training

Explore over 18 Microduck RL locomotion tasks for bipedal robot training, including velocity control, recovery, manipulation, and acrobatics. Train on flat or rough terrain.

deep-dive
Sep 8, 2026
How Backlash Is Simulated in Microduck RL Environments: A Technical Deep-Dive

Learn how backlash is simulated in Microduck RL environments. Discover the three-layer system for modeling mechanical play, actuator behavior, and task configurations for accurate reinforcement learning.

deep-dive
Sep 8, 2026

Have a question about this repo?

These articles cover the highlights, but your codebase questions are specific. Give your agent direct access to the source. Share this with your agent to get started:

Share the following with your agent to get started:
curl -s "https://instagit.com/install.md"

Works with
Claude Codex Cursor VS Code OpenClaw Any MCP Client

Maintain an open-source project? Get it listed too →