Stars
CAPO: Critic-Guided Action-Aligned Policy Optimization for Advancing LLM Agent Capabilities
SGLang is a high-performance serving framework for large language models and multimodal models.
An Open-Source Large-Scale Reinforcement Learning Project for Search Agents
An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models
A version of verl to support diverse tool use [TMLR 2026]
本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)
Pytorch code for EMNLP 2023 accepted-main paper "How to Enhance Causal Discrimination of Utterances: A Case on Affective Reasoning" and paper "Learning a Structural Causal Model for Intuition Reaso…