Skip to content
View bchao1's full-sized avatar
🚶‍♂️
I need to focus.
🚶‍♂️
I need to focus.

Block or report bchao1

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion

Python 1,068 61 Updated Jul 22, 2026

SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer

Python 9,143 731 Updated Sep 21, 2026

Solve puzzles. Improve your pytorch.

Jupyter Notebook 4,339 401 Updated Jul 15, 2024

A skill to stop your coding agent from burying the answer. ADHD-friendly output.

Python 50,978 2,942 Updated Sep 19, 2026

Official implementation of MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models

Python 44 2 Updated Jun 12, 2026

Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Python 1,524 192 Updated Aug 20, 2026

Official PyTorch Implementation of Unified Video Action Model (RSS 2025)

Python 413 34 Updated Aug 21, 2026

[ICML 2026] Official implementation for "DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion".

Python 360 41 Updated May 18, 2026

[to appear at NeurIPS 2026] Official implementation of "MilliVid: Adaptive Latents for Long-Range Consistency in Video Generation"

Python 113 13 Updated Sep 24, 2026

Sekai2: From World Exploration to Interactive World Modeling

Python 73 Updated Sep 9, 2026

A unified multimodal model toolkit

Python 634 147 Updated Sep 12, 2026

[ICLR & NeurIPS 2025] Repository for Show-o series, One Single Transformer to Unify Multimodal Understanding and Generation.

Python 1,977 94 Updated Jan 8, 2026

An open-source AI agent that brings the power of Gemini directly into your terminal.

TypeScript 107,153 14,629 Updated Sep 24, 2026

[ICLR 2026] Official Implementation of Muddit [Meissonic II]: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model.

Python 122 2 Updated Apr 13, 2026

MMaDA - Open-Sourced Multimodal Large Diffusion Language Models (dLLMs with block diffusion, mixed-CoT, unified RL)

Python 1,674 90 Updated Feb 14, 2026

NVIDIA FastGen: Fast Generation from Diffusion Models

Python 1,011 84 Updated Aug 21, 2026

🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking like a caveman.

Go 107,712 6,239 Updated Sep 24, 2026

Official implementation of "Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer" (ECCV 2026)

Python 5 Updated Jun 21, 2026

The official code for NeurIPS 2025 "MagCache: Fast Video Generation with Magnitude-Aware Cache"

Python 279 7 Updated Nov 17, 2025

LLM Wiki is a cross-platform desktop application that turns your documents into an organized, interlinked knowledge base — automatically. Instead of traditional RAG (retrieve-and-answer from scratc…

TypeScript 19,960 2,259 Updated Aug 25, 2026

Florence-2 is a novel vision foundation model with a unified, prompt-based representation for a variety of computer vision and vision-language tasks.

Jupyter Notebook 223 20 Updated Jul 3, 2024

Gemma open-weight LLM library, from Google DeepMind

Python 5,746 1,027 Updated Sep 16, 2026
Python 120 Updated Jun 12, 2026

Code repository for "Spectral Progressive Diffusion for Efficient Image and Video Generation"

Python 23 3 Updated Aug 13, 2026

Code release for "Foveated Diffusion: Efficient Spatially Adaptive Image and Video Generation"

Python 21 3 Updated Sep 1, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 2 Updated Sep 21, 2026

An agentic skills framework & software development methodology that works.

Shell 291,215 26,053 Updated Sep 22, 2026

Ideogram 4: Open image model at the forefront of design

Python 2,851 288 Updated Jun 30, 2026

Wrapper of 50+ image matching models with a unified interface

Python 916 82 Updated Sep 7, 2026

Implementation of Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players

Python 662 14 Updated Jun 17, 2026
Next