Starred repositories
A RLHF Infrastructure for Vision-Language Models
Robust recipes to align language models with human and AI preferences
[NeurIPS 2024] SimPO: Simple Preference Optimization with a Reference-Free Reward
Repository for Meta Chameleon, a mixed-modal early-fusion foundation model from FAIR.
🔥🔥 LLaVA++: Extending LLaVA with Phi-3 and LLaMA-3 (LLaVA LLaMA-3, LLaVA Phi-3)
Recipes to train reward model for RLHF.
NeurIPS 2024 Paper: A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing
[CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. 接近GPT-4o表现的开源多模态对话模型
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
An Open-source Toolkit for LLM Development
Official repo for "Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models"
LLaVA-Plus: Large Language and Vision Assistants that Plug and Learn to Use Skills
Syntax Error-Free and Generalizable Tool Use for LLMs via Finite-State Decoding
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
RLHF implementation details of OAI's 2019 codebase
RL algorithm: Advantage induced policy alignment
【TMM 2025🔥】 Mixture-of-Experts for Large Vision-Language Models
Fine-tune LLM agents with online reinforcement learning
This repo is meant to serve as a guide for Machine Learning/AI technical interviews.
LLM (Large Language Model) FineTuning
PyTorch Implementation of "V* : Guided Visual Search as a Core Mechanism in Multimodal LLMs"
Machine Learning Engineering Open Book
Aligning pretrained language models with instruction data generated by themselves.
A generalized information-seeking agent system with Large Language Models (LLMs).
Data and code for the Corr2Cause paper (ICLR 2024)