Skip to content
View kongzhecn's full-sized avatar

Block or report kongzhecn

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Distill Minimax-H3 into 4 steps

Python 234 9 Updated Aug 15, 2026

An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System

Python 22,978 2,786 Updated Aug 13, 2026

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

Python 35,737 4,092 Updated Aug 12, 2026

Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

Python 22,781 2,628 Updated May 25, 2026
Python 416 46 Updated Jul 17, 2026
Python 6,050 368 Updated Aug 15, 2026

MAGI-2-preview: Scaling Video Generation Models Efficiently

Python 534 14 Updated Aug 6, 2026

[ICML2026] Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization

Python 61 8 Updated Jul 26, 2026

Official repo for paper "Echo-Infinity: Learnable Evolving Memory for Real-Time Infinite Video Generation"

Python 106 3 Updated Jun 4, 2026

Official code for Echo-Memory: a controlled study of memory in action-conditioned video world models (Context, Compression, Spatial, State-Space).

Python 232 15 Updated Aug 16, 2026

Code for the paper HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement

Python 175 12 Updated Aug 9, 2026

[SIGGRAPH Asia 26 Conditionally Accept]PAct: Part-Decomposed Single-View Articulated Object Generation

Jupyter Notebook 77 3 Updated Jul 21, 2026

VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo

Python 2,152 251 Updated Aug 14, 2026

On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators

Python 547 44 Updated Aug 10, 2026

Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence

Python 920 45 Updated Aug 5, 2026

Infinite Worlds with Versatile Interactions

Python 1,519 110 Updated Jul 14, 2026

[ECCV 2026] WildWorld: A Large-Scale Dataset for Dynamic World Modeling with Actions and Explicit State toward Generative ARPG

Python 379 5 Updated Jul 20, 2026

Code for MIRA: Multiplayer Interactive World Models with Representation Autoencoders

Python 503 30 Updated Aug 3, 2026

EchoStyle: Unlocking High-Fidelity Video Stylization with Reverse Data Synthesis

Python 32 2 Updated Jul 2, 2026

DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation

Python 166 16 Updated Jun 26, 2026

AI PPT赛道终结者,史上最最最强 PPT Skill!!! 使用GPT生成豪华的图片格式PPT,然后转换为完全可编辑的PPTX文件。

Python 1,769 154 Updated Jun 7, 2026

Wan: Open and Advanced Large-Scale Video Generative Models

Python 17,149 2,167 Updated Mar 17, 2026
Python 11,885 817 Updated Feb 9, 2026

A toolkit for speaker diarization.

Jupyter Notebook 524 62 Updated Aug 4, 2026

Official Implementation of LongLive-RAG: A general retrieval-augmented framework for long video generation.

Python 107 2 Updated Jun 4, 2026

JoyAI-Echo: Pushing the Frontier of Long Audio-Visual Generation

Python 1,857 168 Updated Jun 26, 2026

Official page of ImmerIris: A Large-Scale Dataset and Benchmark for Off-Axis and Unconstrained Iris Recognition in Immersive Applications.

HTML 30 1 Updated Jun 9, 2026

"CLI-Anything: Making ALL Software Agent-Native" -- CLI-Hub: https://clianything.cc/

Python 47,620 4,411 Updated Aug 13, 2026

Implementation of Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players

Python 654 13 Updated Jun 17, 2026

A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models

Python 780 22 Updated Jun 15, 2026
Next