Skip to content
View zzhanghub's full-sized avatar
🛸
At work
🛸
At work

Organizations

@graphic-design-ai

Block or report zzhanghub

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[ECCV 2026] The official implementation of paper "PixelU: A U-Shaped Transformer for Efficient End-to-End Pixel Diffusion"

25 1 Updated Jul 6, 2026

[ECCV 2026] Official implementation of Hi-DiT: Hybrid Latent-Pixel Diffusion Transformer for Image Generation

Python 7 Updated Jun 22, 2026

Official JAX code of MiniT2I.

Python 125 4 Updated Jun 19, 2026

A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.

Python 1,311 94 Updated Jul 14, 2026

Ideogram 4: Open image model at the forefront of design

Python 2,697 278 Updated Jun 30, 2026
Python 69 2 Updated May 14, 2026

Cheers: Decoupling Patch Details from Semantic Representations Enables Unified Multimodal Comprehension and Generation

Python 257 21 Updated Apr 13, 2026

SenseNova-U series: Native Unified Paradigm with NEO-unify from the First Principles

Python 4,571 396 Updated Aug 7, 2026
JavaScript 5,139 317 Updated Aug 3, 2026

[NeurIPS 2025] Official implementation for our paper "Scaling Diffusion Transformers Efficiently via μP".

Python 100 2 Updated Nov 2, 2025

Train the smallest LM you can that fits in 16MB. Best model wins!

Python 5,177 3,296 Updated May 4, 2026

Fast, accurate & comprehensive text measurement & layout

TypeScript 49,744 2,730 Updated Jun 23, 2026

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.

Rust 195,019 109,245 Updated Aug 6, 2026
Python 46 3 Updated Mar 19, 2026

AI agents running research on single-GPU nanochat training automatically

Python 93,463 13,279 Updated Mar 26, 2026

[ICML'26] Code and website for Self-Flow: Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis

Python 718 22 Updated May 23, 2026

NextFlow🚀: Unified Sequential Modeling Activates Multimodal Understanding and Generation

331 15 Updated Jan 9, 2026

Qwen-Image-Layered: Layered Decomposition for Inherent Editablity

Python 2,055 165 Updated Dec 31, 2025

A research-friendly PyTorch Lightning toolkit for training, fine-tuning, and evaluating AutoencoderKL for Stable Diffusion and FLUX.

Python 90 9 Updated Apr 20, 2026

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 15,090 1,581 Updated Aug 9, 2026

[ICLR 2026] Official repo of paper "Reconstruction Alignment Improves Unified Multimodal Models". Unlocking the Massive Zero-shot Potential in Unified Multimodal Models through Self-supervised Lear…

Python 411 17 Updated Aug 3, 2026

Ming - facilitating advanced multimodal understanding and generation capabilities built upon the Ling LLM.

Jupyter Notebook 666 59 Updated Jul 27, 2026

[NeurIPS 2025] Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations

Python 202 7 Updated Sep 18, 2025

OneCAT: Decoder-Only Auto-Regressive Model for Unified Understanding and Generation

Python 261 6 Updated Sep 22, 2025

[Fully open] [Encoder-free MLLM] Vision as LoRA

Python 389 31 Updated Jun 12, 2025

Awesome Layout Generation

85 10 Updated Apr 10, 2025

The official repository for LaCTok:Latent Consistency Tokenizer for High-resolution Image Reconstruction and Generation by 256 Tokens

Python 2 Updated Aug 27, 2025

Enjoy the magic of Diffusion models!

Python 12,887 1,264 Updated Aug 7, 2026

VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo

Python 2,135 247 Updated Aug 7, 2026

Open-source SOTA multi-image editing model

Python 871 43 Updated Jul 13, 2026
Next