Skip to content
View vasgaowei's full-sized avatar

Block or report vasgaowei

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[Official Repo] JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion

Python 1,307 44 Updated Aug 15, 2026

Official code for FutureBridge-OPD

Python 19 Updated Jul 31, 2026

[Tech Report] Context Scaling: Scaling Properties of Text Conditioning in Visual Generation

Python 157 2 Updated Aug 7, 2026

MAGI-2-preview: Scaling Video Generation Models Efficiently

Python 520 12 Updated Aug 6, 2026

[ICML 2026] Multimodal deep-research MLLM and benchmark. The first long-horizon multimodal deep-research MLLM, extending the number of reasoning turns to dozens and the number of search-engine inte…

Python 675 56 Updated Aug 8, 2026

[Tech Report] Democratizing the Training of Video World Models from Scratch. 🔥 🔥 🔥

Python 89 6 Updated Aug 4, 2026

Official repository for the ACM MM 2026 paper “Remember-R1: Mitigating Long-Context Visual Forgetting through Reinforcement Learning”

Python 174 3 Updated Aug 4, 2026

[ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation" & Causal Forcing++

Python 920 52 Updated Jul 23, 2026

Official codebase for "Self Forcing: Bridging Training and Inference in Autoregressive Video Diffusion" (NeurIPS 2025 Spotlight)

Python 3,478 281 Updated Sep 12, 2025

DistillAlign: Coordinating Mode Covering and Mode Seeking in Autoregressive Video Distillation

Python 117 4 Updated Jul 30, 2026

Vision-OPD is a regional-to-global on-policy self-distillation framework that transfers a model's own privileged crop-conditioned perception to its full-image policy, enabling fine-grained visual u…

Python 275 11 Updated Jul 17, 2026
Python 1,416 159 Updated Aug 10, 2026

VQRAE: Representation Quantization Autoencoders for Multimodal Understanding, Generation and Reconstruction

Python 14 Updated Mar 26, 2026

iFSQ & LlamaGen-REPA

Python 105 10 Updated Jan 27, 2026

Give Claude the ability to watch any video. /watch downloads, extracts frames, transcribes, hands it all to Claude.

Python 15,511 1,497 Updated Jul 1, 2026

NVIDIA FastGen: Fast Generation from Diffusion Models

Python 943 80 Updated Aug 4, 2026
Python 10 2 Updated Jul 21, 2026

[arXiv 2026] This is the official PyTorch implementation of "MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators".

Python 81 Updated Jul 18, 2026

Perceive-to-Reason: Decoupling Perception and Reasoning for Fine-Grained Visual Reasoning

Python 34 1 Updated Jul 8, 2026

[ECCV2026] Official Implementation of "VideoSearch-R1: Iterative Video Retrieval and Reasoning via Soft Query Refinement"

Python 33 1 Updated Jul 2, 2026

TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs

Python 119 1 Updated Jul 27, 2026

AI-Powered Agentic Repository Intelligence Platform

Python 1 Updated Jul 20, 2026

Boogu-Image-0.1 is an Apache-2.0 open-source image generation and editing model family that delivers near-closed-source performance with an order of magnitude less data.

Python 949 56 Updated Jul 23, 2026

Sutskever 30 implementations inspired by https://papercode.vercel.app/ | For Agents, use https://github.com/pageman/Sutskever-Agent | Polyglot / Multi-Backed version at https://github.com/pageman/s…

Jupyter Notebook 4,248 555 Updated Mar 15, 2026
Python 910 87 Updated Aug 15, 2026
Python 31 Updated May 28, 2026

On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators

Python 543 43 Updated Aug 10, 2026
Next