A lean, ROS-free sim-to-real framework for training and deploying Vision-Language-Action (VLA) models and RL agents. Native MuJoCo Gymnasium wrappers with synchronous execution for Franka, UR5e, xArm, SO101 and YAM.
-
Updated
Sep 5, 2026 - Python
A lean, ROS-free sim-to-real framework for training and deploying Vision-Language-Action (VLA) models and RL agents. Native MuJoCo Gymnasium wrappers with synchronous execution for Franka, UR5e, xArm, SO101 and YAM.
Generative Unified Instruction-based Demonstration Environment for robotic data collection and manipulation.
OmniPatrol-VLA: 基于 Qwen2-VL 与 ROS2 的分层具身决策巡检系统。支持 LoRA 微调与 4-bit 量化,实现端云协同的智慧交通违章自动研判。 (English: A multi-modal smart traffic inspection robot powered by Qwen2-VL and ROS2. Features LoRA fine-tuning, 4-bit quantization, and edge-cloud hierarchical architecture.)
A curated collection of resources for Physical AI — papers, frameworks, datasets, simulators, and projects.
An asynchronous, neuro-symbolic VLA (Vision-Language-Action) orchestration stack for edge autonomy. Fuses probabilistic Qwen2-VL visual reasoning and faster-whisper ASR with deterministic PX4/MAVSDK flight-control loops and HSV color guardrails.
Heterogeneous Edge SoC & SystemC simulation platform for real-time SmolVLA Vision-Language-Action (VLA) robotics
To associate your repository with the vla-models topic, visit your repo's landing page and select "manage topics."