-
in beijing
- china
- not yet
Lists (3)
Sort Name ascending (A-Z)
Starred repositories
The batteries-included agent harness.
Pre-indexed code knowledge graph, auto syncs on code changes, for Claude Code, Codex, Gemini, Cursor, OpenCode, AntiGravity, Kiro, and Hermes Agent — fewer tokens, fewer tool calls, 100% local
Textbook on reinforcement learning from human feedback
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
Ongoing research training transformer models at scale
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …
AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI
Free, open-source ASO keyword research tool — self-hosted via Docker. No API keys, no accounts, no data leaves your machine.
An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of…
Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
Automatically redacts sensitive data in screenshots before sending to AI agents
If you can read ~100 lines of Python, you understand agents.
The simplest, fastest repository for training/finetuning medium-sized GPTs.
Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
Algorithm powering the For You feed on X
🦔 PostHog is the leading platform for building self-driving products. Our developer tools – AI observability, analytics, session replay, flags, experiments, error tracking, logs, and more – capture…
Scalable, fast, and lightweight system for large-scale topic modeling
Multi-thread implementation of Factorization Machines with FTRL for binary-class classification problem.
Multi-thread implementation of lambdaFM with FTRL for ranking problem. LambdaFM is a learning-to-rank algorithm by combining LambdaRank and Factorization Machines.
A library for efficient similarity search and clustering of dense vectors.
A framework for large scale recommendation algorithms.
A high-throughput and memory-efficient inference and serving engine for LLMs
A course on aligning smol models.
Train transformer language models with reinforcement learning.
FastGPT is a knowledge-based platform built on the LLMs, offers a comprehensive suite of out-of-the-box capabilities such as data processing, RAG retrieval, and visual AI workflow orchestration, le…
RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs