Skip to content
View alvinliu0's full-sized avatar

Block or report alvinliu0

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Awesome Visual Tokenizers/Autoencoders

20 Updated Nov 19, 2025

[IEEE TPAMI 2026] Simulating the Real World: Survey & Resources, which contains our survey "Simulating the Real World: A Unified Survey of Multimodal Generative Models" (IEEE TPAMI, 2026) and Aweso…

383 13 Updated Aug 13, 2026

[CVPR 2025] HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation

Python 63 5 Updated Jul 8, 2025

Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the future state of the world in the form of video.

Python 1,346 190 Updated Jun 8, 2026

Cosmos-Predict2 is a collection of general-purpose world foundation models for Physical AI that can be fine-tuned into customized world models for downstream applications.

Python 793 105 Updated Oct 29, 2025

[ICLR’26] Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Control

Python 108 2 Updated Feb 8, 2026

Cosmos-Predict1 is a collection of general-purpose world foundation models for Physical AI that can be fine-tuned into customized world models for downstream applications.

Jupyter Notebook 465 83 Updated Jun 7, 2026

Cosmos-Transfer1 is a world-to-world transfer model designed to bridge the perceptual divide between simulated and real-world environments.

Python 815 106 Updated Jun 7, 2026

NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

Jupyter Notebook 11,510 827 Updated Aug 14, 2026

[ICLR'25] 3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation

Jupyter Notebook 371 18 Updated Jul 4, 2025

[ICLR 2025] EdgeRunner: Auto-regressive Auto-encoder for Efficient Mesh Generation

Python 311 16 Updated Dec 22, 2024

[ICLR 2025] Implementation of Accelerating Auto-regressive Text-to-Image Generation with Training-free Speculative Jacobi Decoding

Python 53 4 Updated Apr 21, 2025

A Video Tokenizer Evaluation Dataset

Python 158 14 Updated Jan 13, 2025

A suite of image and video neural tokenizers

Jupyter Notebook 1,733 92 Updated Feb 11, 2025

TC4D: Trajectory-Conditioned Text-to-4D Generation

Python 205 6 Updated Oct 15, 2024

[ECCV 2024] The official implementation of paper "BrushNet: A Plug-and-Play Image Inpainting Model with Decomposed Dual-Branch Diffusion"

Python 1,743 146 Updated Dec 17, 2024

[ICCV 2023] The official implementation of paper "HumanSD: A Native Skeleton-Guided Diffusion Model for Human Image Generation"

Python 305 21 Updated Oct 24, 2023

4D-fy: Text-to-4D Generation Using Hybrid Score Distillation Sampling

Python 340 10 Updated Dec 10, 2024

[CVPR 2024 Highlight] Code for "HumanGaussian: Text-Driven 3D Human Generation with Gaussian Splatting"

Python 493 44 Updated Dec 30, 2023

[ICLR 2024] Github Repo for "HyperHuman: Hyper-Realistic Human Generation with Latent Structural Diffusion"

HTML 495 13 Updated Oct 14, 2023

[CVPR'2023] Taming Diffusion Models for Audio-Driven Co-Speech Gesture Generation

Python 265 19 Updated Mar 18, 2026

The repository for paper Unsupervised Volumetric Animation

Python 69 1 Updated Sep 22, 2023

Elucidating the Design Space of Diffusion-Based Generative Models (EDM)

Python 1,989 211 Updated Mar 16, 2024

Text-to-3D & Image-to-3D & Mesh Exportation with NeRF + Diffusion.

Python 8,855 768 Updated Dec 10, 2023

Machine Learning Interviews from FAANG, Snapchat, LinkedIn. I have offers from Snapchat, Coupang, Stitchfix etc. Blog: mlengineer.io.

12,781 2,042 Updated Aug 31, 2023

Official PyTorch Implementation of "Learning to Learn with Generative Models of Neural Network Checkpoints"

Python 347 23 Updated Oct 3, 2022

Official implementation of Cold-Diffusion for different transformations in pytorch.

Python 1,136 83 Updated Oct 13, 2022

Awesome Lists for Tenure-Track Assistant Professors and PhD students. (助理教授/博士生生存指南)

Python 1,642 97 Updated Feb 1, 2024

[CVPR 2022] Code for "Learning Hierarchical Cross-Modal Association for Co-Speech Gesture Generation"

Python 144 9 Updated Mar 16, 2023

Tackling the Generative Learning Trilemma with Denoising Diffusion GANs https://arxiv.org/abs/2112.07804

Python 759 89 Updated Dec 2, 2022
Next