Here are
212 public repositories
matching this topic...
Mano-P: Open-source GUI-VLA agent for edge devices. #1 on OSWorld (specialized, 58.2%). Runs locally on Apple M4 Mac mini/MacBook — no data leaves your device.Mano-P 是一个开源 GUI-VLA 项目,支持在 Mac mini/MacBook 上或通过算力棒本地运行推理,实现纯视觉驱动的跨平台 GUI 自动化操作。数据完全本地处理,支持复杂多步骤任务规划与执行。
NVIDIA Alpamayo 1 Nano is an open 10B reasoning VLA model for autonomous vehicles that pairs driving trajectories with Chain-of-Causation reasoning.
Updated
Sep 9, 2026
Python
[CVPR 2025] Open-source, End-to-end, Vision-Language-Action model for GUI Agent & Computer Use.
Updated
Apr 24, 2026
Python
A curated collection of papers, explainers, and resources on World Action Models for embodied AI
[NeurIPS 2025] AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning
Updated
May 29, 2026
Python
本项目旨在为致力于进入VLA(Vision-Language-Action)领域的算法工程师提供一份全中文、实战导向的学习/面试手册。 不同于通用的 CV/NLP 面试指南,本项目聚焦于 Robotics 特有的挑战
Updated
Sep 12, 2026
HTML
[ICLR 2026] ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
Updated
Sep 11, 2026
Python
[ICLR 2026] Towards Unified Latent VLA for Whole-body Loco-manipulation Control
An Open-World Foundation Model for General-Purpose Embodied Intelligence.
Updated
Sep 10, 2026
Python
🌐 Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future
Open source evals for physical AI. Run any LLM/VLA on any arm/humanoid against any real/sim benchmark.
Updated
Sep 2, 2026
Python
a minimal, beginner-friendly VLA to show how robot policies can fuse images, text, and states to generate actions
Updated
Mar 17, 2026
Python
NVIDIA Alpamayo 1.5 Nano is an open 10B reasoning VLA model for autonomous vehicles with reinforcement-learning enhanced reasoning, navigation guidance, and visual question answering.
Updated
Sep 9, 2026
Python
A unified framework for training, fine-tuning, and evaluating World Action Models
Updated
Sep 12, 2026
Python
[CVPR 2025] The offical Implementation of "Universal Actions for Enhanced Embodied Foundation Models"
Updated
Nov 6, 2025
Python
Curated embodied AI list: surveys, VLA models, datasets, simulators, humanoids, robot learning, and safety resources.
Updated
Sep 7, 2026
Python
NVIDIA Alpamayo 2 Super is an open 34B multi-task foundation model designed to supercharge autonomous vehicle development.
Updated
Sep 9, 2026
Python
✨✨Official implementation of BridgeVLA and BridgeVLA++
Updated
Aug 13, 2026
Python
Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations
Updated
Jul 16, 2026
Python
🔥 Data Pyramid for Embodied Manipulation: A Survey
Add this topic to your repo
To associate your repository with the
vision-language-action
topic, visit your repo's landing page and select "manage topics."
Learn more
You can’t perform that action at this time.