Skip to content
View kongzhecn's full-sized avatar

Block or report kongzhecn

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[ICML2026] Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization

Python 61 6 Updated Jul 26, 2026

Official repo for paper "Echo-Infinity: Learnable Evolving Memory for Real-Time Infinite Video Generation"

Python 105 3 Updated Jun 4, 2026

A Simple Baseline for Video World Models with Memory

Python 239 17 Updated Jul 29, 2026

Code for the paper HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement

Python 158 12 Updated Jul 21, 2026

[SIGGRAPH Asia 26 Conditionally Accept]PAct: Part-Decomposed Single-View Articulated Object Generation

Jupyter Notebook 71 2 Updated Jul 21, 2026

VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo

Python 2,112 239 Updated Jul 29, 2026

On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators

Python 456 21 Updated Jul 22, 2026

Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence

Python 880 40 Updated Jul 10, 2026

Infinite Worlds with Versatile Interactions

Python 1,424 97 Updated Jul 14, 2026

[ECCV 2026] WildWorld: A Large-Scale Dataset for Dynamic World Modeling with Actions and Explicit State toward Generative ARPG

Python 419 5 Updated Jul 20, 2026

Code for MIRA: Multiplayer Interactive World Models with Representation Autoencoders

Python 465 24 Updated Jul 12, 2026

EchoStyle: Unlocking High-Fidelity Video Stylization with Reverse Data Synthesis

Python 31 1 Updated Jul 2, 2026

DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation

Python 164 16 Updated Jun 26, 2026

AI PPT赛道终结者,史上最最最强 PPT Skill!!! 使用GPT生成豪华的图片格式PPT,然后转换为完全可编辑的PPTX文件。

Python 1,656 143 Updated Jun 7, 2026

Wan: Open and Advanced Large-Scale Video Generative Models

Python 16,884 2,112 Updated Mar 17, 2026
Python 11,796 808 Updated Feb 9, 2026

A toolkit for speaker diarization.

Jupyter Notebook 506 60 Updated May 29, 2026

Official Implementation of LongLive-RAG: A general retrieval-augmented framework for long video generation.

Python 101 2 Updated Jun 4, 2026

JoyAI-Echo: Pushing the Frontier of Long Audio-Visual Generation

Python 1,838 165 Updated Jun 26, 2026

Official page of ImmerIris: A Large-Scale Dataset and Benchmark for Off-Axis and Unconstrained Iris Recognition in Immersive Applications.

HTML 30 1 Updated Jun 9, 2026

"CLI-Anything: Making ALL Software Agent-Native" -- CLI-Hub: https://clianything.cc/

Python 46,265 4,310 Updated Jul 9, 2026

Implementation of Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players

Python 650 11 Updated Jun 17, 2026

A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models

Python 748 20 Updated Jun 15, 2026

Codex skill for converting slide images, PDFs, and image-based PPTX files into editable PowerPoint decks.

Python 1,668 84 Updated Jul 28, 2026

Multimodal RL training framework for diffusion & omni models

Python 674 112 Updated Jul 29, 2026

Official Repo of "D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models"

Python 291 8 Updated May 22, 2026

A unified framework for easy reinforcement learning in Flow-Matching models

Python 643 51 Updated Jul 12, 2026

AI turns documents or topics into real, native PowerPoint decks—with native shapes, transitions and animations, data-backed charts and tables on demand, audio narration from speaker notes, and supp…

Python 41,731 3,458 Updated Jul 29, 2026

Interactive World Model papers organized by core research challenges.

Python 276 9 Updated Jul 16, 2026

[ICML 2026] World-R1: Reinforcing 3D Constraints for Text-to-Video Generation

Python 411 16 Updated Jun 3, 2026
Next