vla
Tracked open-source repos tagged vla, sorted by stars.
Related topics
Topics that frequently appear alongside vla on the same repo.
Recent risers
Repos created in the last 90 days, tagged vla.
- #1
🐕 MindPaw — 基于 ESP8266 的开源桌面四足机器狗,支持语音、手势、网页遥控、豆包 AI 对话与 PAD 情感交互。可选 AI Infra 网关,适合机器人入门、嵌入式 AI、边缘 Agent 与 HRI 研究。
★ 2,463+21Star change over the last 7 days - #2
NVIDIA Alpamayo 1 Nano is an open 10B reasoning VLA model for autonomous vehicles that pairs driving trajectories with Chain-of-Causation reasoning.
★ 2,013+6Star change over the last 7 days - #3
[ICLR 2026] SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
★ 1,844+6Star change over the last 7 days - #4
A Pragmatic VLA Foundation Model
★ 1,794+11Star change over the last 7 days - #5★ 1,416+9Star change over the last 7 days
- #6★ 1,126+1Star change over the last 7 days
- #7
InternRobotics' open platform for building generalized navigation foundation models.
★ 1,090+18Star change over the last 7 days - #8
Unified Codebase for Advanced World Models.
★ 866+2Star change over the last 7 days - #9
[NeurIPS 2025 spotlight] Official implementation for "FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving"
★ 831+9Star change over the last 7 days - #10
🚀🚀🚀A collection of some awesome public projects about Large Language Model(LLM), Vision Language Model(VLM), Vision Language Action(VLA), AI Generated Content(AIGC), the related Datasets and Applications.
★ 815+1Star change over the last 7 days - #11
🔥 SpatialVLA: a spatial-enhanced vision-language-action model that is trained on 1.1 Million real robot episodes. Accepted at RSS 2025.
★ 718+1Star change over the last 7 days - #12
Running VLA at 30Hz frame rate and 480Hz trajectory frequency
★ 611+2Star change over the last 7 days - #13
本项目旨在为致力于进入VLA(Vision-Language-Action)领域的算法工程师提供一份全中文、实战导向的学习/面试手册。 不同于通用的 CV/NLP 面试指南,本项目聚焦于 Robotics 特有的挑战
★ 594+33Star change over the last 7 days - #14
[ICLR2026] Official implementation for "JanusVLN: Decoupling Semantics and Spatiality with Dual Implicit Memory for Vision-Language Navigation"
★ 588+4Star change over the last 7 days - #15
FlashRT is a high-performance realtime inference engine for small-batch, latency-sensitive AI workloads. The flagship integration is production VLA control for Pi0, Pi0.5, GROOT N1.6, and Pi0-FAST. Also support llm e.g, qwen3.6-27B
★ 550+11Star change over the last 7 days - #16
A high-performance framework for training LLMs, VLMs, diffusion, and embodied models on NVIDIA GPUs and Kunlun XPUs.
★ 549+15Star change over the last 7 days - #17
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
★ 529+0Star change over the last 7 days