仅需Python基础,从0构建自己的具身智能机器人;从0逐步构建VLA/OpenVLA/SmolVLA/Pi0, 深入理解具身智能
-
Updated
Sep 8, 2026 - Python
仅需Python基础,从0构建自己的具身智能机器人;从0逐步构建VLA/OpenVLA/SmolVLA/Pi0, 深入理解具身智能
Robotics Data Toolkit | Convert between robotics dataset formats (RLDS, LeRobot v2/v3, Zarr, HDF5, Rosbag). Inspect, visualize, and analyze datasets. Works with HuggingFace Hub. Built for OpenVLA, Octo, LeRobot, and Diffusion Policy workflows.
PickAgent: OpenVLA-powered Pick and Place Agent | Gradio&Simulation | Vision Language Action Model
ROS2-native runtime, benchmark, and adapter hub for Vision-Language-Action models.
Independent VLA research notes: OpenVLA / π0 / Spirit paper reviews, LIBERO reproduction, XLeRobot integration. Transitioning from AD motion planning to embodied AI.
Using VLM-based visual question answering to perceive scenes and control robots in MuJoCo simulation.
Train and deploy Vision-Language-Action models natively on Apple Silicon.
Think Less, Act Early: Reinforced Latent Reasoning with Early Exit in Vision-Language-Action Models
FR3 robot in Gazebo integrated with 4-bit quantized OpenVLA and MoveIt
A full-stack Embodied AI simulation suite powered by Genesis World: featuring OpenVLA closed-loop evaluation, procedural scene generation, massively parallel GPU RL (Unitree Go2), and multi-physics coupling (PBD/SPH).
A RoboNix Skill for experience-memory retrieval and verified action reuse across OpenVLA and π0.
Safety monitors for learned robot policies under distribution shift. A benchmark for which monitor still works once the deployment distribution moves.
VLA manipulation ablations, held-out generalization, and an OpenVLA-7B feasibility study, all measured on a 6GB laptop GPU
Fine-tuning OpenVLA on 53,294 Isaac Sim observations — contact improves, target selection degrades
OpenVLA-compatible action-model layer + model-agnostic robot agent runtime (state machine, safety validation, failure recovery), verified in tabletop sim, LIBERO and Gazebo/ROS 2
Unofficial PyTorch reproduction for OpenVLA: An Open-Source Vision-Language-Action Model.
Simulated UR5e + MoveIt 2 cell for testing vision-language-action policies in ROS 2 Jazzy. Swap SmolVLA, OpenVLA-7B (4-bit) or GR00T N1.7 behind one ZeroMQ protocol — all runnable on a 6 GB GPU.
Minecraft → VLA robot interface. Expose a Minecraft player as a standard embodied AI chassis (camera, joint states, action space) for VLA models like SmolVLA, OpenVLA, π₀.₅.
To associate your repository with the openvla topic, visit your repo's landing page and select "manage topics."