Skip to content
View dddd-d's full-sized avatar

Block or report dddd-d

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

CAPO: Critic-Guided Action-Aligned Policy Optimization for Advancing LLM Agent Capabilities

Python 23 1 Updated Aug 11, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,761 7,869 Updated Aug 14, 2026

An Open-Source Large-Scale Reinforcement Learning Project for Search Agents

Python 606 39 Updated Nov 26, 2025

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,361 305 Updated Aug 14, 2026

A version of verl to support diverse tool use [TMLR 2026]

Python 1,031 88 Updated Jul 15, 2026

青稞Talk

218 2 Updated Jul 27, 2026

大厂发布的AI落地实践、顶尖实验室的最新论文、工业界的真实踩坑记录

3,690 511 Updated Jun 9, 2026

本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)

HTML 24,896 2,841 Updated Jul 19, 2026

Pytorch code for EMNLP 2023 accepted-main paper "How to Enhance Causal Discrimination of Utterances: A Case on Affective Reasoning" and paper "Learning a Structural Causal Model for Intuition Reaso…

Python 18 2 Updated Aug 20, 2024