Skip to content
View luow39's full-sized avatar

Highlights

  • Pro

Block or report luow39

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

An easy-to-use, fast toolkit to scale up RL post-training on a single node.

Python 290 120 Updated Aug 18, 2026

Official Implementation of OCR-free Document Understanding Transformer (Donut) and Synthetic Document Generator (SynthDoG), ECCV 2022

Python 6,913 565 Updated Jul 11, 2024

DocLayNet: A Large Human-Annotated Dataset for Document-Layout Analysis

453 26 Updated Feb 1, 2023

TAT-QA (Tabular And Textual dataset for Question Answering) contains 16,552 questions associated with 2,757 hybrid contexts from real-world financial reports.

Python 138 28 Updated Dec 9, 2024

Skills for Real Engineers. Straight from my .agents directory.

Shell 221,396 19,071 Updated Aug 17, 2026

The code and resource of "Towards Comprehensive Detection of Chinese Harmful Memes" (NeurIPS2024 D&B).

Python 86 3 Updated May 17, 2025

[AAAI'25 (Oral)] Jailbreaking Large Vision-language Models via Typographic Visual Prompts

Python 212 13 Updated Jun 26, 2025

[ACL 2025] Can We Trust AI Doctors? A Survey of Medical Hallucination in Large Language and Large Vision-Language Models

12 Updated Aug 25, 2025

up-to-date curated list of state-of-the-art Large vision language models hallucinations research work, papers & resources

327 19 Updated Feb 8, 2026

One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks

Python 4,367 640 Updated Aug 6, 2026

An agentic skills framework & software development methodology that works.

Shell 273,631 24,482 Updated Aug 13, 2026

Tensor library for machine learning

C++ 15,193 1,782 Updated Aug 18, 2026

Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++

C++ 6,788 740 Updated Aug 12, 2026
Python 11,903 817 Updated Feb 9, 2026

A toy PyTorch implementation of FLUX diffusion transformers

Python 14 2 Updated Aug 3, 2026

斯坦福CS146S 现代软件开发者(vibe coding)课程中文版。 本中文课程由RapidAI 赞助。英文版: https://themodernsoftware.dev/

193 21 Updated Dec 26, 2025

[ICLR'26] SPEED: Scalable, Precise, and Efficient Concept Erasure for Diffusion Models

Python 42 1 Updated Mar 9, 2026

Algorithm powering the For You feed on X

Rust 31,937 5,243 Updated Aug 18, 2026

Official implementation of NeurIPS'24 paper "Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models". This work adversarially unlearns the text encoder to enh…

Jupyter Notebook 55 3 Updated Nov 4, 2024

Image inpainting tool powered by SOTA AI Model. Remove any unwanted object, defect, people from your pictures or erase and replace(powered by stable diffusion) any thing on your pictures.

Python 23,352 2,514 Updated Apr 29, 2025

Official Repo for Paper "OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision" [ICLR2025]

145 2 Updated Jan 27, 2025

The image prompt adapter is designed to enable a pretrained text-to-image diffusion model to generate images with image prompt.

Jupyter Notebook 6,669 427 Updated Jun 28, 2024

The Open-Source Multimodal AI Agent Stack: Connecting Cutting-Edge AI Models and Agent Infra

TypeScript 38,633 3,897 Updated Aug 5, 2026
Python 37 3 Updated Feb 13, 2024

T2I-Adapter

Python 3,805 226 Updated Jun 21, 2024

[ICCV 2023] Consistent Image Synthesis and Editing

Python 842 36 Updated Aug 19, 2024

Separable Diffusion Model Unlearning

Python 13 Updated Jan 29, 2025

Officail Implementation for "ReNoise: Real Image Inversion Through Iterative Noising"

Python 266 12 Updated Jul 3, 2024
Next