Skip to content
View thuwzy's full-sized avatar
🎃
Focusing
🎃
Focusing

Block or report thuwzy

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models

Python 738 19 Updated Jun 15, 2026

[ECCV 2026] Official code of GEM: Generative Supervision Helps Embodied Intelligence

Python 89 1 Updated May 30, 2026

[ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation" & Causal Forcing++

Python 879 52 Updated Jul 23, 2026

Advancing Open-source World Models

Python 4,276 391 Updated Jul 9, 2026

VIGA: Vision-as-Inverse-Graphics Agent

Python 1,257 126 Updated May 6, 2026

Muon is an optimizer for hidden layers in neural networks

Python 2,731 131 Updated May 24, 2026

SAM 3D Objects

Python 7,167 855 Updated Jun 2, 2026
C++ 239 11 Updated Jul 24, 2026

[ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation

Python 697 42 Updated Nov 20, 2025

Qwen-Image-Lightning: Speed up Qwen-Image model with distillation

Python 1,340 45 Updated Jan 1, 2026

Resources and paper list for "Thinking with Images for LVLMs". This repository accompanies our survey on how LVLMs can leverage visual information for complex reasoning, planning, and generation.

1,493 47 Updated Mar 9, 2026

Directly Aligning the Full Diffusion Trajectory with Fine-Grained Human Preference

Python 1,278 42 Updated May 11, 2026

Official code of RDT 2

Python 795 57 Updated Feb 7, 2026

Hunyuan3D-Omni: A Unified Framework for Controllable Generation of 3D Assets

Python 598 56 Updated Oct 17, 2025

ViPE: Video Pose Engine for Geometric 3D Perception

Python 2,051 166 Updated Jun 9, 2026

A curated collection of fun and creative examples generated with Nano Banana & Nano Banana Pro🍌, Gemini-2.5-flash-image based model. We also release Nano-consistent-150K openly to support the commu…

23,324 2,376 Updated Dec 12, 2025

Voyager is an interactive RGBD video generation model conditioned on camera input, and supports real-time 3D reconstruction.

Python 1,573 164 Updated Apr 15, 2026
Jupyter Notebook 500 29 Updated Dec 8, 2025

4DNeX: Feed-Forward 4D Generative Modeling Made Easy

Python 840 14 Updated Dec 14, 2025

Hunyuan-GameCraft: High-dynamic Interactive Game Video Generation with Hybrid History Condition

Python 732 75 Updated Nov 28, 2025

Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory

Python 2,277 248 Updated Mar 30, 2026

Generate large-scale explorable 3D scenes with high-quality panorama videos from a single image or text prompt.

Python 771 58 Updated Nov 25, 2025

Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.

Python 8,165 517 Updated Feb 10, 2026

[NeurIPS 2024 & TPAMI 2026] Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers

Python 216 13 Updated Apr 12, 2026

Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels with Hunyuan3D World Model

Python 2,881 262 Updated Apr 15, 2026

PhysX: Physical-Grounded 3D Asset Generation (NeurIPS 2025, Spotlight)

Jupyter Notebook 382 21 Updated Dec 18, 2025

[CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer

Python 13,989 1,520 Updated May 19, 2026

Towards a Generative 3D World Engine for Embodied Intelligence

Python 564 30 Updated Jul 17, 2026

[ICLR'26] Topology-Preserved Auto-regressive Mesh Generation in the Manner of Weaving Silk

Python 117 4 Updated Mar 2, 2026
Next