Stars
A digital data-generation pipeline that synthesizes humanoid loco-manipulation data from 3D assets and video priors.
💻 vibe coding 2026 | Your First Modern Coding course beginners to master step by step.
PhyAgentOS is a self-evolving embodied AI operating system built on agentic workflows.
[RSS 2026] The first framework enabling humanoid robots to learn whole-body loco-manipulation from egocentric human demos
[RSS 2026] Code for RISE: Self-Improving Robot Policy with Compositional World Model
[CVPR 2026] UniDex: A Robot Foundation Suite for Universal Dexterous Hand Control from Egocentric Human Videos
Welcome to GR00T Whole-Body Control (WBC)! This is a unified platform for developing and deploying advanced humanoid controllers. This includes: Decoupled WBC models used in NVIDIA Isaac-Gr00t, Gr0…
[CVPR 2026] 3D Motion Reconstruction for 4D Synthesis
[ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling
MotionGPT3: Human Motion as a Second Modality, a MoT-based framework for unified motion understanding and generation
[CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
[ICCV2025] LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
Official implementation for WorldScore: A Unified Evaluation Benchmark for World Generation
Enjoy the magic of Diffusion models!
Wan: Open and Advanced Large-Scale Video Generative Models
jonstephens85 / Cosmos-WSL2
Forked from NVIDIA/cosmosCosmos is a world model development platform that consists of world foundation models, tokenizers and video processing pipeline to accelerate the development of Physical AI at Robotics & AV labs. C…
SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
CVPR and NeurIPS poster examples and templates
Text2Place : Affordance aware Semantic Mask Generation
[CVPR 2024 Oral] Rethinking Inductive Biases for Surface Normal Estimation
Industry leading face manipulation platform
Codes for the CVPR 2024 paper: "KITRO: Refining Human Mesh by 2D Clues and Kinematic-tree Rotation"
GenZI: Zero-Shot 3D Human-Scene Interaction Generation (CVPR 2024)
[Arxiv-2024] MotionLLM: Understanding Human Behaviors from Human Motions and Videos
MambaOut: Do We Really Need Mamba for Vision? (CVPR 2025)
Awesome work on hand pose estimation/tracking