-
Snap Inc.
- Mountain View, California
- wangkua1.github.io
- @kcjacksonwang
- in/jackson-wang-97a066139
- kcjacksonwang
Stars
[ICLR 2026] UniEdit-Flow: Unleashing Inversion and Editing in the Era of Flow Models
A curated list of papers, code, and resources pertaining to generative image composition or object insertion.
Official repo of On Exact Inversion of DPM-Solvers by Hong et al, in CVPR 2024.
phymhan / prompt-to-prompt
Forked from google/prompt-to-promptOfficial implementation of the paper "ProSpect: Prompt Spectrum for Attribute-Aware Personalization of Diffusion Models"(SIGGRAPH Asia 2023)
VideoSys: An easy and efficient system for video generation
[AAAI 2023] Exploring CLIP for Assessing the Look and Feel of Images
🔎 🖼️ 🔥PyTorch Toolbox for Image Quality Assessment, including PSNR, SSIM, LPIPS, FID, NIQE, NRQM(Ma), MUSIQ, TOPIQ, NIMA, DBCNN, BRISQUE, PI and more...
Codes for ID-Specific Video Customized Diffusion
A collection of resources on controllable generation with text-to-image diffusion models.
LAVIS - A One-stop Library for Language-Vision Intelligence
(CVPR 2024) 🧩 TokenCompose: Text-to-Image Diffusion with Token-level Supervision
Official Implementation of 'Inserting Anybody in Diffusion Models via Celeb Basis'
Subject-Diffusion:Open Domain Personalized Text-to-Image Generation without Test-time Fine-tuning
Official implementations for paper: Anydoor: zero-shot object-level image customization
[ECCV 2024] Viewpoint Textual Inversion: Discovering Scene Representations and 3D View Control in 2D Diffusion Models
[CVPR 2024] 4K4D: Real-Time 4D View Synthesis at 4K Resolution
Demo code for VIRD - VR badminton match video analysis tool
[CVPR 2023] Official Pytorch code for PROB: Probabilistic Objectness for Open World Object Detection
Source code for the paper: "AutoDecoding Latent 3D Diffusion Models"
[CVPR 2023] Code for "Learning Neural Volumetric Representations of Dynamic Humans in Minutes"
Official Code and Dataset for "High-fidelity 3D Human Digitization from Single 2K Resolution Images" (CVPR 2023 Highlight)
This is an official implementation of our CVPR 2023 paper "Human Pose as Compositional Tokens" (https://arxiv.org/pdf/2303.11638.pdf)
[NeurIPS 2023] Official Pytorch code for LOVM: Language-Only Vision Model Selection
4DHumans: Reconstructing and Tracking Humans with Transformers