Skip to content
View DesmonDay's full-sized avatar

Block or report DesmonDay

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 898 85 Updated Aug 6, 2026

Evaluation harness for Apodex-1.0 on public deep-research benchmarks.

Python 380 38 Updated Jun 8, 2026

large language model internal-medicine monitor toolbox

Python 14 10 Updated Aug 8, 2026

agentUniverse is a LLM multi-agent framework that allows developers to easily build multi-agent applications.

Python 2,322 419 Updated Jul 28, 2026

Structured deep research skill for Claude Code/Open Code/Codex with human-in-the-loop control

Python 1,927 156 Updated May 7, 2026

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

Python 1,951 352 Updated Aug 11, 2026

Framework for evaluating and improving agents

Python 4,098 1,516 Updated Aug 11, 2026

An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

Python 566 137 Updated Aug 10, 2026

Code and implementations for the paper "AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning" by Zhiheng Xi et al.

Python 842 84 Updated Feb 15, 2026

Runnable ClaudeCode source code

TypeScript 3,277 3,822 Updated Apr 8, 2026

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.

Rust 195,027 109,220 Updated Aug 6, 2026

Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflo…

Python 140,972 22,657 Updated Aug 10, 2026

Democratizing Reinforcement Learning for LLMs

Python 5,775 599 Updated Aug 11, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,839 1,130 Updated Aug 11, 2026

A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.

Python 4,722 780 Updated May 17, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,655 7,790 Updated Aug 11, 2026

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

C++ 6,229 1,064 Updated Aug 11, 2026

Nano vLLM

Python 14,942 2,443 Updated Apr 26, 2026

[ICLR 2026]QeRL enables RL for 32B LLMs on a single H100 GPU.

Python 513 52 Updated Mar 30, 2026

An elegant \LaTeX\ résumé template. 大陆镜像 https://gods.coding.net/p/resume/git

TeX 11,326 2,870 Updated Mar 15, 2024

Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team.

Python 16,777 1,237 Updated Mar 24, 2026

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 15,108 1,585 Updated Aug 11, 2026

This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."

MATLAB 17,428 1,659 Updated Aug 10, 2026

High-performance safetensors model loader

Python 163 34 Updated Aug 7, 2026

DeepEP: an efficient expert-parallel communication library

Cuda 9,973 1,371 Updated Aug 5, 2026
Python 615 75 Updated Sep 23, 2025

个人中文简历 Latex 源码 https://hijiangtao.github.io/

TeX 3,479 718 Updated Sep 4, 2024

📄 适合中文的简历模板收集(LaTeX,HTML/JS and so on)由 @hoochanlon 维护

8,153 523 Updated Jul 22, 2026

A bidirectional pipeline parallelism algorithm for computation-communication overlap in DeepSeek V3/R1 training.

Python 2,989 327 Updated Jan 14, 2026
Next