Skip to content
View xumwen's full-sized avatar
🏠
Working from home
🏠
Working from home
  • Peking University
  • Beijing

Block or report xumwen

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

AxisRL is an agentic RL post-training framework built on SGLang rollout, Megatron training, and real-world agent workflows.

Python 1,038 22 Updated Aug 3, 2026

AxisAgentic: An Extensible Runtime and Trajectory-Collection Framework for Long-Horizon Agents.

Python 1,103 140 Updated Jul 24, 2026
TypeScript 12 Updated Jul 22, 2026

A Claude Code plugin that shows what's happening - context usage, active tools, running agents, and todo progress

JavaScript 27,246 1,265 Updated Aug 7, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,356 304 Updated Aug 9, 2026

🏆 Top-1 on 5+ benchmarks | Web UI | Supports MiroThinker, Claude, Kimi, OpenAI

Python 3,090 322 Updated Jul 6, 2026

OpenSeeker: A search agent with open-source data and models

Python 765 60 Updated Jun 25, 2026
Python 4,596 503 Updated Apr 22, 2026

🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN

Python 77,555 8,024 Updated Jul 30, 2026

Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL

Python 5,278 477 Updated Nov 13, 2025

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 385,690 81,067 Updated Aug 9, 2026

MiroThinker is a deep research agent optimized for complex research and prediction tasks. Our latest models, MiroThinker-1.7, achieves 74.0 and 75.3 on the BrowseComp and BrowseComp Zh, respectively.

Python 8,369 643 Updated Jul 6, 2026

[TMLR 2024] Official implementation of "NuTime: Numerically Multi-Scaled Embedding for Large-Scale Time-Series Pretraining".

Python 12 1 Updated Dec 23, 2024

An elegant \LaTeX\ résumé template. 大陆镜像 https://gods.coding.net/p/resume/git

TeX 11,324 2,870 Updated Mar 15, 2024

A Survey of Reinforcement Learning for Large Reasoning Models

TeX 2,476 131 Updated Aug 1, 2026

The official Python library for the OpenAI API

Python 31,325 5,115 Updated Aug 7, 2026

gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI

Python 20,296 2,133 Updated Jul 24, 2026

Official repository for the paper "LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code"

Python 926 197 Updated Jul 16, 2025

Research and development (R&D) is crucial for the enhancement of industrial productivity, especially in the AI era, where the core aspects of R&D are mainly focused on data and models. We are commi…

Python 14,183 1,821 Updated Aug 4, 2026

An Open-source RL System from ByteDance Seed and Tsinghua AIR

Python 1,851 85 Updated May 11, 2025

The rule-based evaluation subset and code implementation of Omni-MATH

Python 29 2 Updated Dec 23, 2024

repo for paper https://arxiv.org/abs/2504.13837

Python 347 20 Updated Jul 27, 2026

All-in-one text de-duplication

Python 765 79 Updated Mar 9, 2026

[ICLR'25] BigCodeBench: Benchmarking Code Generation Towards AGI

Python 519 74 Updated Jan 3, 2026

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

Python 42,888 4,925 Updated Aug 9, 2026

Fast and memory-efficient exact attention

Python 24,657 2,973 Updated Aug 9, 2026

Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends

Python 2,511 525 Updated Jun 29, 2026

Fully open reproduction of DeepSeek-R1

Python 26,432 2,446 Updated Apr 2, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,585 7,763 Updated Aug 9, 2026

Unleashing the Power of Reinforcement Learning for Math and Code Reasoners

Python 740 44 Updated Jun 6, 2025
Next