Starred repositories
open-source agentic AI data assistant for the next generation of AI + Data products.
mPLUG-Owl: The Powerful Multi-modal Large Language Model Family
Curated tutorials and resources for Large Language Models, Text2SQL, Text2DSL、Text2API、Text2Vis and more.
LayerDiffuse in pure diffusers without any GUI
A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone
Lora beYond Conventional methods, Other Rank adaptation Implementations for Stable diffusion.
mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding
[AAAI 2025] Official implementation of "OOTDiffusion: Outfitting Fusion based Latent Diffusion for Controllable Virtual Try-on"
[AAAI 2024] AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
Official implementations for paper: Zero-shot Image Editing with Reference Imitation
MambaMorph: a Mamba-based Framework for Medical MR-CT Deformable Registration
Implementation of 💍 Ring Attention, from Liu et al. at Berkeley AI, in Pytorch
Large World Model -- Modeling Text and Video with Millions Context
Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
Official PyTorch implementation of SynDiff described in the paper (https://arxiv.org/abs/2207.08208).
Automated dense category annotation engine that serves as the initial semantic labeling for the Segment Anything dataset (SA-1B).
Staged Training for Transformer Language Models
ResShift: Efficient Diffusion Model for Image Super-resolution by Residual Shifting (NeurIPS@2023 Spotlight, TPAMI@2024)
PDF GPT allows you to chat with the contents of your PDF file by using GPT capabilities. The most effective open source solution to turn your pdf files in a chatbot!
AI PDF chatbot agent built with LangChain & LangGraph
Repo for BenCao [original name: HuaTuo (华驼)], Instruction-tuning Large Language Models with Chinese Medical Knowledge. 本草(原名:华驼)模型仓库,基于中文医学知识的大语言模型指令微调
BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.
[ECCV 2022] XMem: Long-Term Video Object Segmentation with an Atkinson-Shiffrin Memory Model
Making large AI models cheaper, faster and more accessible