Skip to content
View azshue's full-sized avatar

Organizations

@judy-vscode

Block or report azshue

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.

Python 2,771 227 Updated Jul 24, 2026

My learning notes for ML SYS.

HTML 6,890 482 Updated Aug 18, 2026

Cambrian-1 is a family of multimodal LLMs with a vision-centric design.

Python 2,013 139 Updated Nov 7, 2025

Solve puzzles. Improve your pytorch.

Jupyter Notebook 4,290 393 Updated Jul 15, 2024

Official Repo for Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning

Jupyter Notebook 417 41 Updated Dec 15, 2024

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 23,026 4,425 Updated Aug 19, 2026

Minimal reproduction of DeepSeek R1-Zero

Python 13,223 1,578 Updated Feb 27, 2026

[EMNLP-2024] Build multimodal language agents for fast prototype and production

Python 2,664 292 Updated Mar 19, 2025

AllenAI's post-training codebase

Python 3,833 578 Updated Aug 19, 2026

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,933 1,002 Updated Aug 13, 2026
Python 70 3 Updated Jun 2, 2026

A instruction data generation system for multimodal language models.

Jupyter Notebook 37 1 Updated Jan 31, 2025

Extend existing LLMs way beyond the original training length with constant memory usage, without retraining

Python 735 44 Updated Apr 10, 2024

🍃 MINT-1T: A one trillion token multimodal interleaved dataset.

832 19 Updated Jul 31, 2024

[COLM-2024] List Items One by One: A New Data Source and Learning Paradigm for Multimodal LLMs

Python 147 4 Updated Aug 23, 2024

Official repo for Detecting, Explaining, and Mitigating Memorization in Diffusion Models (ICLR 2024)

Python 80 9 Updated Apr 3, 2024

Package to optimize Adversarial Attacks against (Large) Language Models with Varied Objectives

Python 70 6 Updated Feb 22, 2024

An open-source framework for training large multimodal models.

Python 4,119 319 Updated Aug 31, 2024

LAVIS - A One-stop Library for Language-Vision Intelligence

Jupyter Notebook 11,263 1,109 Updated Jun 2, 2026

A Next-Generation Training Engine Built for Ultra-Large MoE Models

Python 5,179 444 Updated Aug 19, 2026

Consistency Distilled Diff VAE

Python 2,213 81 Updated Nov 7, 2023

The simplest, fastest repository for training/finetuning medium-sized GPTs.

Python 62,198 10,725 Updated Nov 12, 2025

Ongoing research training transformer models at scale

Python 17,473 4,374 Updated Aug 19, 2026

Official repository of NEFTune: Noisy Embeddings Improves Instruction Finetuning

Python 412 19 Updated May 17, 2024

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

Python 42,957 4,936 Updated Aug 17, 2026

Denoising Diffusion Implicit Models

Python 1,844 232 Updated Jul 26, 2024

GLIDE: a diffusion-based text-conditional image synthesis model

Python 3,686 499 Updated Mar 8, 2024

An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.

Python 39,511 4,783 Updated May 1, 2026

A framework for few-shot evaluation of language models.

Python 13,711 3,498 Updated Aug 14, 2026
Next