-
Nvidia
- Shanghai
-
18:28
(UTC +08:00) - https://dorniwang.github.io/
Stars
[CVPR 2026] Humanoid-GPT: Scaling Data and Structure for Zero-Shot Motion Tracking
[CVPR 2026] Towards Real-Time Diffusion-Based Streaming Video Super-Resolution — An efficient one-step diffusion framework for streaming VSR with locality-constrained sparse attention and a tiny co…
Infinite Worlds with Versatile Interactions
An Open-World Foundation Model for General-Purpose Embodied Intelligence.
An Open-Source World Model for Action-Conditioned Embodied Intelligence.
FlashRT is a high-performance realtime inference engine for small-batch, latency-sensitive AI workloads. The flagship integration is production VLA control for Pi0, Pi0.5, GROOT N1.6, and Pi0-FAST.…
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation
Agent Skills for NVIDIA products — install into Claude Code, Codex, and other coding agents to run Physical AI, robotics, simulation, CUDA, and RAG workflows end to end.
PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image (CVPR 2026)
AI-powered agents for automating 3D content workflows using Vision-Language Models (VLMs). Content Agents analyze 3D assets and automate material assignment, physics property classification, and te…
Our inference and training framework to run on the Cosmos Models
PyTorch code and models for VJEPA2 self-supervised learning from video.
Isaac Lab API, powered by MuJoCo-Warp, for RL and robotics research
UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms
Building General-Purpose Robots Based on Embodied Foundation Model
[ECCV 2026] VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model
Collect some World Models for Autonomous Driving (and Robotic, etc.) papers.
[AAAI 2026 Oral] FreeAskWorld is an interactive simulation framework that integrates large language models (LLMs) for high-level planning and socially grounded interaction in embodied AI.
Momentum Human Rig is an anatomically-inspired parametric full-body digital human model developed at Meta. It includes: A parametric body skeletal model; A realistic 3D mesh skinned to the skeleton…
SimWorld: An Open-ended Realistic Simulator for Autonomous Agents in Physical and Social Worlds
Official implementation of Tuna-2: Pixel Embeddings Beat Vision Encoders for Unified Understanding and Generation
An Agentic System for Scalable Articulated 3D Asset Generation
The Causal Nexus for Embodied AI. A physical commonsense firewall and Knowledge-as-a-Service (KaaS) platform.
HumanNet: Scaling Human-centric Video Learning to One Million Hours
Sample Environment for the LeRobot SO-101 Robot in Isaac Lab to collect demonstrations in a simulation