Skip to content
View wshenx's full-sized avatar

Block or report wshenx

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

FireRed-OpenStoryline is an AI video editing agent that transforms manual editing into intention-driven directing through natural language interaction, LLM-powered planning, and precise tool orches…

Python 3,162 371 Updated May 7, 2026

Establishing new state-of-the-art results for Bokeh Rendering on the EBB! Dataset.

Python 17 2 Updated Aug 25, 2023

This is a sample C# project that extracts Depth and Color information from videos shot in iPhone's Cinematic mode and outputs each as separate videos, along with a sample Unity project for 3D playb…

C# 48 6 Updated Oct 6, 2023

[NeurIPS'25 Spotlight] Official repository for "Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference Alignment"

Python 777 76 Updated Sep 27, 2025

Scaling Diffusion Transformers with Mixture of Experts

Python 435 20 Updated Sep 9, 2024

Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation

Python 1,961 97 Updated Aug 15, 2024

Janus-Series: Unified Multimodal Understanding and Generation Models

Python 17,756 2,235 Updated Feb 1, 2025

[ICLR & NeurIPS 2025] Repository for Show-o series, One Single Transformer to Unify Multimodal Understanding and Generation.

Python 1,965 94 Updated Jan 8, 2026
Python 272 13 Updated Jul 23, 2024
Jupyter Notebook 25 Updated Aug 29, 2024

[CVPR 2025] Diffusion-4K: Ultra-High-Resolution Image Synthesis with Latent Diffusion Models

Python 364 15 Updated Nov 24, 2025
Python 2,231 157 Updated Nov 8, 2024

OmniGen: Unified Image Generation. https://arxiv.org/pdf/2409.11340

Jupyter Notebook 4,335 362 Updated Dec 4, 2025

[ICLR 2025] Official implementation of Posterior-Mean Rectified Flow: Towards Minimum MSE Photo-Realistic Image Restoration

Python 753 42 Updated Feb 5, 2025

Official Implementation and Dataset of "PPR10K: A Large-Scale Portrait Photo Retouching Dataset with Human-Region Mask and Group-Level Consistency", CVPR 2021

Python 346 24 Updated Jan 14, 2025

[CVPR 2025] CoDe: Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient

Python 108 5 Updated Sep 27, 2025

[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". A…

Jupyter Notebook 8,714 571 Updated Nov 10, 2025

[ECCV 2024 Oral 🔥] Arc2Face: A Foundation Model for ID-Consistent Human Faces ------------------------ [ICCVW 2025] ID-Consistent, Precise Expression Generation with Blendshape-Guided Diffusion

Python 803 62 Updated Oct 10, 2025
Python 47 8 Updated Aug 4, 2024

A set of nodes for ComfyUI that can composite layer and mask to achieve Photoshop like functionality.

Python 3,112 200 Updated Jul 23, 2026

More relighting!

Python 8,477 522 Updated Feb 20, 2025

cocos效果

TypeScript 8 5 Updated May 3, 2024

Smooth random noise generator

Python 74 3 Updated Mar 27, 2026

Inference code for Llama models

Python 59,525 9,800 Updated Jan 26, 2025

[CVPR2024] Official code for Drag Your Noise: Interactive Point-based Editing via Diffusion Semantic Propagation

Python 87 5 Updated Apr 16, 2024
Python 3 Updated May 10, 2024

[ECCV 2024 Oral] EDTalk - Official PyTorch Implementation

Python 468 39 Updated Sep 29, 2025

Official implementation of the CVPR 2024 paper "FSRT: Facial Scene Representation Transformer for Face Reenactment from Factorized Appearance, Head-pose, and Facial Expression Features"

Python 125 12 Updated Oct 28, 2025
Next