Skip to content
View pku-wuwei's full-sized avatar
🎯
Focusing
🎯
Focusing
  • Peking University
  • Beijing

Block or report pku-wuwei

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Language model tokenization at GB/s

Rust 4,000 211 Updated Aug 6, 2026

🚀 Ultra Recipe for Training Long-Horizon Search Agents - matching frontier AI's search capability with a 20B model + stateful harness

Python 981 153 Updated Jun 15, 2026

feishu-cli 是一个功能完整的飞书开放平台命令行工具。它将飞书文档、知识库、电子表格、消息、日历、任务等操作封装为简洁的命令行接口,核心能力是 Markdown ↔ 飞书文档双向无损转换。

Go 1,352 140 Updated Jul 31, 2026

PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.

Java 28,415 2,709 Updated Aug 13, 2026

Heuristic Learning Blog Post

Python 610 60 Updated May 25, 2026

The batteries-included agent harness.

Python 27,792 3,881 Updated Aug 15, 2026

Multimodal Agentic Document QA benchmark (MADQA)

Python 41 1 Updated Mar 13, 2026

Nano vLLM

Python 15,007 2,462 Updated Apr 26, 2026

A benchmark for evaluating AI agents on frontier ultra long-horizon auto research tasks.

Python 159 18 Updated Jun 17, 2026

A unified library for building, evaluating, and storing speculative decoding algorithms for LLM inference in vLLM

Python 730 183 Updated Aug 15, 2026

Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflo…

Python 141,555 22,722 Updated Aug 14, 2026

A fast and soft pattern search for trillion-scale corpora.

Python 239 11 Updated Feb 28, 2026

AgentIR is a retriever specialized for Deep Research agents.

Python 62 6 Updated Apr 16, 2026

A Heterogeneous Benchmark for Information Retrieval. Easy to use, evaluate your models across 15+ diverse IR datasets.

Python 2,269 250 Updated Oct 16, 2025

BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent (ACL 2026 Main)

Python 332 57 Updated May 28, 2026

AI agents running research on single-GPU nanochat training automatically

Python 93,897 13,308 Updated Mar 26, 2026

a recursive self-improving harness designed to help your agents (and future iterations of those agents) succeed on any task

Python 1,272 108 Updated Aug 15, 2026

PyTorch building blocks for the OLMo ecosystem

Python 1,471 304 Updated Aug 15, 2026

This API provides programmatic access to the AlphaGenome model developed by Google DeepMind.

Python 1,972 271 Updated Jul 31, 2026

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

Python 1,999 361 Updated Aug 15, 2026

A Reproduction of GDM's Nested Learning Paper

Python 708 101 Updated Feb 25, 2026

😊 TPTT: Transforming Pretrained Transformers into Titans

Python 65 2 Updated Jun 7, 2026

Quick illustration of how one can easily read books together with LLMs. It's great and I highly recommend it.

Python 3,832 487 Updated Nov 18, 2025

Kimi Code CLI is your next CLI agent.

Python 11,185 1,291 Updated Aug 3, 2026
Python 4,600 503 Updated Apr 22, 2026

[R]einforcement [L]earning from [M]odel-rewarded [T]hinking - code for the paper "Language Models That Think, Chat Better"

Python 129 7 Updated Oct 27, 2025

🌾 OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.

Python 671 63 Updated Jan 29, 2026
Next