Dexbotic: Open-Source Vision-Language-Action Toolbox
-
Updated
Sep 16, 2026 - Python
Dexbotic: Open-Source Vision-Language-Action Toolbox
An Open-World Foundation Model for General-Purpose Embodied Intelligence.
NVIDIA Alpamayo 1 Nano is an open 10B reasoning VLA model for autonomous vehicles that pairs driving trajectories with Chain-of-Causation reasoning.
[RSS 2025] Learning to Act Anywhere with Task-centric Latent Actions
InternRobotics' open platform for building generalized navigation foundation models.
Unified Codebase for Advanced World Models.
[NeurIPS 2025 spotlight] Official implementation for "FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving"
🚀🚀🚀A collection of some awesome public projects about Large Language Model(LLM), Vision Language Model(VLM), Vision Language Action(VLA), AI Generated Content(AIGC), the related Datasets and Applications.
🔥 SpatialVLA: a spatial-enhanced vision-language-action model that is trained on 1.1 Million real robot episodes. Accepted at RSS 2025.
本项目旨在为致力于进入VLA(Vision-Language-Action)领域的算法工程师提供一份全中文、实战导向的学习/面试手册。 不同于通用的 CV/NLP 面试指南,本项目聚焦于 Robotics 特有的挑战
A high-performance framework for training LLMs, VLMs, diffusion, and embodied models on NVIDIA GPUs and Kunlun XPUs. Supports Pi0.5, GR00T-N1.7, FastWAM, Cosmos3, DreamZero, DeepSeek-V4, GLM-5.3, Kimi-K3 and more.
Open source evals for physical AI. Run any LLM/VLA on any arm/humanoid against any real/sim benchmark.
人形机器人运动智能论文、开源项目、产业与求职知识库
FlashRT is a high-performance realtime inference engine for small-batch, latency-sensitive AI workloads. The flagship integration is production VLA control for Pi0, Pi0.5, GROOT N1.6, and Pi0-FAST. Also support llm e.g, qwen3.6-27B
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
To associate your repository with the vla topic, visit your repo's landing page and select "manage topics."