vla
Here are 467 public repositories matching this topic...
Dexbotic: Open-Source Vision-Language-Action Toolbox
-
Updated
Sep 16, 2026 - Python
An Open-World Foundation Model for General-Purpose Embodied Intelligence.
-
Updated
Sep 16, 2026 - Python
NVIDIA Alpamayo 1 Nano is an open 10B reasoning VLA model for autonomous vehicles that pairs driving trajectories with Chain-of-Causation reasoning.
-
Updated
Sep 9, 2026 - Python
[RSS 2025] Learning to Act Anywhere with Task-centric Latent Actions
-
Updated
Nov 19, 2025 - Python
InternRobotics' open platform for building generalized navigation foundation models.
-
Updated
Mar 10, 2026 - Jupyter Notebook
Unified Codebase for Advanced World Models.
-
Updated
Aug 23, 2026 - Python
[NeurIPS 2025 spotlight] Official implementation for "FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving"
-
Updated
May 8, 2026 - Python
🚀🚀🚀A collection of some awesome public projects about Large Language Model(LLM), Vision Language Model(VLM), Vision Language Action(VLA), AI Generated Content(AIGC), the related Datasets and Applications.
-
Updated
Aug 1, 2025
🔥 SpatialVLA: a spatial-enhanced vision-language-action model that is trained on 1.1 Million real robot episodes. Accepted at RSS 2025.
-
Updated
Jun 23, 2025 - Python
本项目旨在为致力于进入VLA(Vision-Language-Action)领域的算法工程师提供一份全中文、实战导向的学习/面试手册。 不同于通用的 CV/NLP 面试指南,本项目聚焦于 Robotics 特有的挑战
-
Updated
Sep 20, 2026 - HTML
FlashRT is a high-performance realtime inference engine for small-batch, latency-sensitive AI workloads. The flagship integration is production VLA control for Pi0, Pi0.5, GROOT N1.6, and Pi0-FAST. Also support llm e.g, qwen3.6-27B
-
Updated
Sep 19, 2026 - C++
A high-performance framework for training LLMs, VLMs, diffusion, and embodied models on NVIDIA GPUs and Kunlun XPUs.
-
Updated
Sep 18, 2026 - Python
人形机器人运动智能论文、开源项目、产业与求职知识库
-
Updated
Sep 19, 2026
Open source evals for physical AI. Run any LLM/VLA on any arm/humanoid against any real/sim benchmark.
-
Updated
Sep 19, 2026 - Python
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
-
Updated
Sep 8, 2026 - Python
Add this topic to your repo
To associate your repository with the vla topic, visit your repo's landing page and select "manage topics."