Starred repositories
VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo
High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.
飞书官方出品的 OpenClaw 飞书/Lark Channel 插件
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
基于多智能体LLM的中文金融交易框架 - TradingAgents中文增强版
EAFT(Entropy-Adaptive Fine-Tuning: Resolving Confident Conflicts to Mitigate Forgetting) official repo
Codec-Aligned Sparsity as a Foundational Principle for Multimodal Intelligence
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
🚀 Efficient implementations for emerging model architectures
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
Design hardware-friendly model architectures and migrate existing LLMs with minimal performance loss
My learning notes for ML SYS.
Janus-Series: Unified Multimodal Understanding and Generation Models
Scalable RL solution for advanced reasoning of language models
保存微信历史版本
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
Robust recipes to align language models with human and AI preferences
A curated list of reinforcement learning with human feedback resources (continually updated)
✨✨Latest Advances on Multimodal Large Language Models
Train transformer language models with reinforcement learning.
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)