Starred repositories
Define and run multi-container applications with Docker
This is a Chinese translation of the CUDA programming guide
对常见IMU芯片的原理、驱动和数据融合算法整理,以区分某度、某坛上面碎片化严重到影响入坑的乱象。也欢迎PR补充内容。
A unified inference and post-training framework for accelerated video generation.
World Model with Physical Information for Robotic Planning Task
仅需Python基础,从0构建自己的具身智能机器人;从0逐步构建VLA/OpenVLA/SmolVLA/Pi0, 深入理解具身智能
moojink / openvla-oft
Forked from openvla/openvlaFine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
openvla / openvla
Forked from TRI-ML/prismatic-vlmsOpenVLA: An open-source vision-language-action model for robotic manipulation.
TartanAir dataset tools and samples
Open source simulator for autonomous vehicles built on Unreal Engine / Unity, from Microsoft AI & Research
求职/升学/程序员找工作简历模板大全,全网最全的简历模板收集 | Free collection of resume templates for job/education |中文简历|英文简历|程序员简历模板|简历大全|简历模板|工作简历|简历封面
Unfied World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets
Unified Codebase for Advanced World Models.
[CVPR 2025 Highlight] GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control
[CVPR 2026] GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation
WorldGrow: Generating Infinite 3D World [AAAI 2026 Oral]
Codebase for "Beyond Pixel Histories: World Models with Persistent 3D State" (ICML 26)
Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the future state of the world in the form of video.
Code for "FlashWorld: High-quality 3D Scene Generation within Seconds" (ICLR 2026 Oral)
Collect some World Models for Autonomous Driving (and Robotic, etc.) papers.
A comprehensive list of papers for the definition of World Models and using World Models for General Video Generation, Embodied AI, and Autonomous Driving, including papers, codes, and related webs…