Skip to content
View xlwangDev's full-sized avatar

Block or report xlwangDev

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion

Python 1,027 61 Updated Jul 22, 2026

A framework for efficient model inference with omni-modality models

Python 6,097 1,465 Updated Aug 14, 2026

[NeurIPS 2025]

JavaScript 9 Updated May 15, 2026

[NeurIPS 2025] ARGenSeg: Image Segmentation with Autoregressive Image Generation Model

Python 8 Updated May 15, 2026

Powerful menu bar manager for macOS

Swift 29,262 842 Updated Sep 20, 2025

A curated collection of fun and creative examples generated with Nano Banana & Nano Banana Pro🍌, Gemini-2.5-flash-image based model. We also release Nano-consistent-150K openly to support the commu…

23,486 2,388 Updated Dec 12, 2025

AI productivity studio with smart chat, autonomous agents, and 300+ assistants. Unified access to frontier LLMs

TypeScript 50,429 4,776 Updated Aug 14, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,208 81,172 Updated Aug 14, 2026

[ECCV 2026 Oral] RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards.

Python 341 19 Updated Jul 27, 2026

HiMTok: Learning Hierarchical Mask Tokens for Image Segmentation with Large Multimodal Model

Python 98 4 Updated Jul 17, 2025

Edit-R1: Reinforce Image Editing with Diffusion Negative-Aware Finetuning and MLLM Implicit Feedback

Python 295 12 Updated Jan 24, 2026

[ICLR 2026 Oral] DiffusionNFT: Online Diffusion Reinforcement with Forward Process

Python 1,011 44 Updated Feb 10, 2026

A project page template for academic papers. Demo at https://eliahuhorwitz.github.io/Academic-project-page-template/

JavaScript 5,147 1,146 Updated Sep 4, 2025

Code release for Ming-UniVision: Joint Image Understanding and Geneation with a Continuous Unified Tokenizer

Python 143 5 Updated Oct 14, 2025

Some Conferences' accepted paper lists (including AI, ML, Robotic)

Python 1,338 83 Updated Jan 23, 2025

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,361 305 Updated Aug 14, 2026

Extracted system prompts from Anthropic - Claude Fable 5, Opus 5, Claude Design, Claude Code. OpenAI - ChatGPT GPT-5.6-Sol, Codex. Google - Gemini 3.5 Flash, 3.1 Pro, Antigravity. xAI - Grok, Curso…

JavaScript 62,855 10,327 Updated Aug 13, 2026

Implement a ChatGPT-like LLM in PyTorch from scratch, step by step

Jupyter Notebook 102,614 15,725 Updated Aug 10, 2026

Awesome Unified Multimodal Models

1,311 46 Updated Mar 24, 2026

[🚀ICML 2025] "Taming Rectified Flow for Inversion and Editing" Using FLUX and HunyuanVideo for image and video editing!

Python 639 20 Updated May 1, 2025

Ming - facilitating advanced multimodal understanding and generation capabilities built upon the Ling LLM.

Jupyter Notebook 667 59 Updated Jul 27, 2026

Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.

Jupyter Notebook 19,781 1,830 Updated Jan 30, 2026

A curated list of publications on image and video segmentation leveraging Multimodal Large Language Models (MLLMs), highlighting state-of-the-art methods, innovative applications, and key advanceme…

232 6 Updated Jun 28, 2026