Stars
- All languages
- Assembly
- Astro
- C
- C#
- C++
- CSS
- ChucK
- Clojure
- CoffeeScript
- Cuda
- Dart
- Elm
- G-code
- Go
- HTML
- Inform 7
- Inno Setup
- Java
- JavaScript
- Julia
- Jupyter Notebook
- Kotlin
- Lua
- MATLAB
- MLIR
- Makefile
- Nim
- OCaml
- Objective-C
- OpenEdge ABL
- PHP
- PLpgSQL
- Perl
- PureBasic
- Python
- R
- Racket
- ReScript
- Roff
- Ruby
- Rust
- SCSS
- Scala
- ShaderLab
- Shell
- Smarty
- Stan
- Svelte
- Swift
- TSQL
- TeX
- TypeScript
- VBA
- Vue
- Web Ontology Language
- XSLT
- Zig
This repo is the official implementation of "τ0-VLA: a Hierarchical Robot Foundation Model with World-Model-Guided Test-Time Computation".
Flow-SDE GRPO post-training of SmolVLA on LIBERO-Plus: diagnostics, an evaluation audit, and a preregistered null. Includes the preregistration, raw per-episode data, and a runnable reproduction of…
[ECCV 2026] MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE
[CVPR 2026 Highlight] LitePT: Lighter Yet Stronger Point Transformer
The Comprehensive Toolkit for Embodied AI Models
Code for InSight: Self-Guided Skill Acquisition via Steerable VLAs -- insight-vla.github.io/
CoRL 2025 TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models
Official code for "μ0: A Scalable 3D Interaction-Trace World Model"
FastCrest Tether: the OSS edge-to-cloud AI deploy CLI. Optimize, verify, deploy across Jetson, RTX, Apple Silicon, AMD. Hybrid edge-cloud inference with parity certs.
Official PyTorch implementation for "Large Language Diffusion Models"
Source code release for "SERF: Spatiotemporal Environment and Robot Feature Map for Long-Horizon Mobile Manipulation"
(ICML 2026) Any3D-VLA: Enhancing VLA Robustness via Diverse Point Clouds
[arXiv 2026] APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies
[NeurIPS 2025] An official implementation of Flow-GRPO: Training Flow Matching Models via Online RL
Try OpenPI closed loop simulation / scenes generation / 3DGS model scenes in genie_sim
BFM_Zero: A Promptable Behavioral Foundation Model for Humanoid Control Using Unsupervised Reinforcement Learning
From Vision-Language-Action Models to a Real-World Robot Learning Stack
VLA-RFT: Vision-Language-Action Models with Reinforcement Fine-Tuning
Official repository for the project "TraceGen: World Modeling in 3D Trace-Space Enables Learning from Cross-Embodiment Videos" (CVPR'26)
[CoRL 2025] ManiFlow: A General Robot Manipulation Policy via Consistency Flow Training
Official code for "TraceGen: World Modeling in 3D Trace-Space Enables Learning from Cross-Embodiment Videos" (CVPR 2026)
QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models
🔥 Official code repository for "Unlocking Dense Metric Depth Estimation in VLMs"
VL-JEPA Joint Embedding Predictive Architecture for Vision-language (Paper-Derived Implementation)Joint Embedding Predictive Architecture for Vision-language (Paper-Derived Implementation)