Skip to content

Repository files navigation

中文版 | English

MimicLite

MimicLite is an efficient, general humanoid motion-tracking system for training deployable PPO and PPO-ROA policies with competitive tracking quality. Under a matched MuJoCo evaluation, MimicLite improves global root tracking over SONIC while achieving comparable local tracking accuracy. The same policy family supports low-latency Pico-driven teleoperation and highly dynamic motion tracking on a physical Unitree G1.

The technical report is available at mimic-lite.pdf.

Project Repositories

This repository is the project landing page. Training, evaluation, dataset conversion, and deployment instructions are maintained in their respective repositories:

Component Repository Contents
MimicLite EGalahad/mimic-lite Training, evaluation, policy export, task configs, and learning code.
Training framework Agent-3154/active-adaptation Simulation backends, distributed launchers, environments, and shared infrastructure.
Motion data toolkit EGalahad/any4hdmi Motion conversion, validation, visualization, and dataset tooling.
Deployment runtime EGalahad/sim2real ONNX inference, MuJoCo sim2sim, Pico teleoperation, and Unitree G1 deployment.

Released Checkpoints

The public release set now exposes only the latest 16x16384 G1 mixture Huge policies. Training compute is reported as GPU hours on RTX 4090 GPUs.

Policy Actor hidden dimensions Parallel environments Checkpoint GPU hours
MimicLite-PPO [1024, 1024, 1024] 16 × 16384 4234dd57 92.3
MimicLite-ROA [1024, 1024, 1024] 16 × 16384 (train -> adapt -> finetune) 9287d8e0 173.2

Download the deploy ONNX and YAML from the shared sim2real artifacts: MimicLite-PPO and MimicLite-ROA. Older Huge/Base/v1.1 releases are retained only in the Drive archive.

Canonical cross-codebase tracking evaluation

For a fair comparison, we report the motion-lookahead latency required by each policy, defined by its furthest required future-reference frame. All values use the shared 50 Hz reference-motion contract.

Policy MimicLite-PPO MimicLite-ROA BFM-Zero SONIC SONIC-v1.1 SONIC low-latency HoloMotion TeleopIT Humanoid-GPT HEFT TWIST2
Motion-lookahead latency 0.08 s 0.08 s 0.12 s 0.90 s 0.90 s 0.18 s 0.20 s 0.00 s 0.02 s 0.12 s 0.00 s

Training Data

Released training datasets are collected in the any4hdmi Hugging Face collection. The BONES-SEED dataset is the exception: to respect its license and redistribution terms, users obtain it from the original source, while EGalahad/any4hdmi provides only the conversion scripts and processing tools.

Deployment Support

The sim2real runtime provides a modular observation interface that separates policy-specific input construction from the shared deployment runtime. Integrating a policy requires only an observation class and a YAML specification; the inference, simulator, and robot interfaces remain unchanged. This common path supports integrated MuJoCo evaluation and real-robot execution for MimicLite, HEFT, TeleopIT, Humanoid-GPT, BFM-Zero, SONIC, and TWIST2. Policy inference is decoupled from robot I/O through interchangeable MuJoCo and physical Unitree G1 backends.

License

This integration repository is released under GPL-3.0-or-later. Component repositories retain their own histories and license files; verify dataset and component licenses before redistribution.

About

No description, website, or topics provided.

Resources

Stars

169 stars

Watchers

0 watching

Forks

Contributors