Skip to content
View PengyunZhao0720's full-sized avatar

Block or report PengyunZhao0720

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 6 Updated Jul 21, 2026

SimWAM: A Simple World Action Model for End-to-End Autonomous Driving.

Python 41 Updated Aug 11, 2026

Wan: Open and Advanced Large-Scale Video Generative Models

Python 16,802 3,237 Updated Mar 5, 2026

A Minimalist, Batteries-included Repository for Advancing World Model Science.

Python 708 43 Updated Jun 15, 2026

[CVPR 2025] Exploring the Deep Fusion of Large Language Models and Diffusion Transformers for Text-to-Image Synthesis

Python 140 5 Updated May 16, 2025

[CVPR 2026] Visual Geometry Transformer for Autonomous Driving

Python 343 19 Updated Jun 10, 2026

[ICML 2025]"Graph World Model", Tao Feng, Yexin Wu, Guanyu Lin, Jiaxuan You

Python 45 3 Updated Sep 20, 2025

[NeurIPS 2025] Video World Models with Long-term Spatial Memory

Python 75 8 Updated May 11, 2026

[ICML 2025] Official PyTorch Implementation of "History-Guided Video Diffusion"

Python 707 40 Updated Jul 1, 2025

code for "Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion"

Python 1,283 72 Updated Jul 6, 2026

(CVPR 2025) From Slow Bidirectional to Fast Autoregressive Video Diffusion Models

Python 1,420 87 Updated Aug 7, 2025

Official repository for "OMG: Omni-Modal Motion Generation for Generalist Humanoid Control", https://arxiv.org/abs/2606.10340.

Python 105 9 Updated Aug 11, 2026

AI-agent Skill for generating polished HTML slide decks: editorial magazine and Swiss layouts, image prompts, social covers, and a WebGL/low-power presentation runtime.

HTML 23,787 1,707 Updated Aug 7, 2026

[CVPR 2026 Oral] Official implementation for ChordEdit: One-Step Low-Energy Transport for Image Editing

Python 392 18 Updated May 13, 2026
Python 154 14 Updated Mar 30, 2026

Modular, scalable library to train ML models

Python 290 31 Updated Aug 7, 2026

Dense Prediction Transformers

Python 2,336 281 Updated Dec 18, 2024

[CVPR 2026 Oral] VGGT Omega

Python 3,941 267 Updated Jul 15, 2026
Python 870 54 Updated Jul 3, 2026

Building General-Purpose Robots Based on Embodied Foundation Model

Python 1,210 97 Updated Jul 21, 2026

DGGT: Feedforward 4D Reconstruction of Dynamic Driving Scenes using Unposed Images

Python 559 60 Updated Jan 15, 2026

RL-based MPC for Discrete-Time Nonlinear Systems (Python)

Python 119 9 Updated Dec 2, 2025

🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning

Python 26,589 5,367 Updated Aug 11, 2026

📹 A more flexible framework that can generate videos at any resolution and creates videos from images.

Python 2,194 169 Updated Aug 11, 2026

自动驾驶笔记,以解析各模块知识点、整合行业优秀解决方案进行阐述,以帮助自己及有需要的读者;包含深度学习、deeplearning、无人驾驶、BEV、Transformer、ADAS、CVPR、特斯拉AI DAY、大模型、chatgpt等内容.

Shell 817 140 Updated Feb 26, 2026

Pytorch implementation for "DeSiRe-GS: 4D Street Gaussians for Static-Dynamic Decomposition and Surface Reconstruction for Urban Driving Scenes"

Python 243 18 Updated Dec 20, 2024

[NeurIPS 2025] PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation

Python 129 7 Updated Feb 22, 2026

[ICRA 2026] Official implementation of the paper "GSWorld: Closed-Loop Photo-Realistic Simulation Suite for Robotic Manipulation"

Python 215 11 Updated Feb 27, 2026
Next