Highlights
- Pro
Stars
[Science Robotics 2026] Official Implementation of the paper RL-100
Official repository of T-Rex: Tactile-Reactive Dexterous Manipulation
High-Level Control Library for Franka Robots with Python and C++ Support
"OpenHarness: Open Agent Harness with a Built-in Personal Agent--Ohmo!"
[NeurIPS 2025] VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models
[CoRL 2025 Oral] AirExo-2: Scaling up Generalizable Robotic Imitation Learning with Low-Cost Exoskeletons
MLA: A Multisensory Language-Action Model for Multimodal Understanding and Forecasting in Robotic Manipulation
An open-source, full-stack robotic neck for active perception
Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?
🦙 LaMa Image Inpainting, Resolution-robust Large Mask Inpainting with Fourier Convolutions, WACV 2022
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
[CVPR 2026] UniDex: A Robot Foundation Suite for Universal Dexterous Hand Control from Egocentric Human Videos
PyTorch implementation of MAR+DiffLoss https://arxiv.org/abs/2406.11838
[ICLR 2025] Autoregressive Video Generation without Vector Quantization
[RSS 2026] Causal video-action world model for generalist robot control
Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals
Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.
Spirit-v1.5: A Robotic Foundation Model by Spirit AI
Official implementation of the project HuDOR: Bridging the Human to Robot Dexterity Gap through Object-Oriented Rewards project. Website: https://object-rewards.github.io
Geometric Retargeting A Principled, Ultrafast Neural Hand Retargeting Algorithm