Skip to content
View fcchit's full-sized avatar

Block or report fcchit

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.

Python 8,561 1,387 Updated Aug 3, 2026

[ECCV 2026 Oral] DreamID-V: Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Diffusion Transformer

Python 669 92 Updated May 22, 2026

[CVPR 2026] Soul: Breathe Life into Digital Human for High-fidelity Long-term Multimodal Animation

Python 64 1 Updated Dec 16, 2025

OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Python 53 Updated Apr 15, 2026

Official inference repo for FLUX.2 models

Python 2,614 189 Updated Mar 12, 2026

Qwen-Image-Lightning: Speed up Qwen-Image model with distillation

Python 1,351 46 Updated Jan 1, 2026
Python 1,747 203 Updated Nov 15, 2025

Qwen-Image text to image lora trainer

Python 766 69 Updated Dec 16, 2025

Face-MakeUp (SD1.5): Multimodal Facial Prompts for Text-to-Image Generation (ECAI-2025)

Python 26 Updated Jan 19, 2025

[CVPR 2025 Highlight🔥] Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Python 852 46 Updated Apr 14, 2026

Long Video Gen Infrastructure

Python 2,531 242 Updated Aug 7, 2026

HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation

Python 3,225 179 Updated Jun 23, 2026

HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning

Python 1,283 246 Updated Jan 25, 2026

[ICLR 2026] Youtu-GraphRAG: Vertically Unified Agents for Graph Retrieval-Augmented Complex Reasoning

Python 1,238 183 Updated Feb 26, 2026

A fast AI Video Generator for the GPU Poor. Supports Wan 2.1/2.2, LTX-2, Qwen Image, Hunyuan Video, LTX Video and Flux.

Python 7,963 1,216 Updated Aug 10, 2026

​​Unlimited-length talking video generation​​ that supports image-to-video and video-to-video generation

Python 7,609 1,331 Updated May 22, 2026

Phantom-Data: Towards a General Subject-Consistent Video Generation Dataset

120 3 Updated Feb 25, 2026

HunyuanImage-2.1: An Efficient Diffusion Model for High-Resolution (2K) Text-to-Image Generation​

Python 675 53 Updated Oct 14, 2025

[ArXiv 2025] A survey about controllable video generation: This repo is the official awesome of "Controllable video generation: A survey"

764 44 Updated Jul 31, 2026

The minimal opencv for Android, iOS, ARM Linux, Windows, Linux, MacOS, HarmonyOS, WebAssembly, watchOS, tvOS, visionOS

C++ 3,335 461 Updated Jul 12, 2026

Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.

Python 8,227 529 Updated Feb 10, 2026

[CVPR2026 🎉] Stand-In is a lightweight, plug-and-play framework for identity-preserving video generation.

Python 780 51 Updated Aug 10, 2026

Official inference repo for FLUX.1 models

Python 25,894 1,911 Updated Jul 31, 2025

Wan: Open and Advanced Large-Scale Video Generative Models

Python 17,056 2,150 Updated Mar 17, 2026

Enjoy the magic of Diffusion models!

Python 12,913 1,264 Updated Aug 11, 2026

The official code of Yume

Python 681 45 Updated Jan 14, 2026

Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.

Python 27,495 2,030 Updated Jan 9, 2026

[NeurIPS 2025 D&B🔥] OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation

Jupyter Notebook 226 9 Updated May 19, 2026

Phantom: Subject-Consistent Video Generation via Cross-Modal Alignment

Python 1,515 98 Updated Sep 11, 2025
Next