Skip to content
View ZichengDuan's full-sized avatar
👀
Focusing
👀
Focusing

Block or report ZichengDuan

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,324 301 Updated Jul 25, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,652 4,274 Updated Jul 25, 2026

Code for MIRA: Multiplayer Interactive World Models with Representation Autoencoders

Python 455 24 Updated Jul 12, 2026
Python 66 Updated Jul 3, 2026

Our inference and training framework to run on the Cosmos Models

Python 413 90 Updated Jul 23, 2026

A Comprehensive Survey of Interactive Video World Models

212 12 Updated Jul 22, 2026

Official implementation of Chain-of-Action: Trajectory Autoregressive Modeling for Robotic Manipulation. Accepted in NeurIPS 2025.

Python 107 7 Updated Dec 13, 2025

WRBench: camera-controlled generation and diagnostic evaluation of video world models.

Python 18 2 Updated Jul 12, 2026

[Official Code] PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory

Python 43 7 Updated Jun 17, 2026

A feed-forward 3D foundation model for reconstructing scenes from streaming data

Python 15,332 1,592 Updated Jul 23, 2026
Python 805 47 Updated Jul 3, 2026

Official Implementation of MultiWorld: Scalable Multi-Agent Multi-View Video World Models

Python 247 13 Updated May 12, 2026

More relighting!

Python 8,470 524 Updated Feb 20, 2025

From Automated Idea Factory to Realization

Shell 1,326 116 Updated Jul 18, 2026

[NeurIPS 2025 D&B🔥] OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation

Jupyter Notebook 222 9 Updated May 19, 2026

[ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modeling

Python 485 8 Updated Apr 16, 2026

Interactive World Model papers organized by core research challenges.

Python 271 9 Updated Jul 16, 2026

A fast AI Video Generator for the GPU Poor. Supports Wan 2.1/2.2, LTX-2, Qwen Image, Hunyuan Video, LTX Video and Flux.

Python 6,720 1,020 Updated Jul 24, 2026

Official code, models, and data for Vista4D: Video Reshooting with 4D Point Clouds (CVPR 2026 Highlight)

Python 560 44 Updated Jun 2, 2026

[CVPR 2026 Highlight] VideoCoF: Unified Video Editing with Temporal Reasoner

Python 204 13 Updated Jun 17, 2026

Karabiner-Elements is a powerful tool for customizing keyboards on macOS

C++ 22,536 918 Updated Jul 25, 2026

The first open-domain closed-loop revisited benchmark for evaluating memory consistency and action control in world models.

Python 74 2 Updated Jul 2, 2026

Next-Token Prediction is All You Need

Python 2,433 100 Updated Jan 12, 2026

Native Multimodal Models are World Learners

Python 1,537 69 Updated Dec 30, 2025

Official repository from the paper "Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind"

Python 17 1 Updated Mar 18, 2025
Python 6 Updated Mar 31, 2026

(ACM MM 2025) Let Your Video Listen to Your Music! – Beat-Aligned,Content-Preserving Video Editing with Arbitrary Music

Python 6 Updated Apr 12, 2026

Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models

Python 267 14 Updated Jul 23, 2026
Next