Skip to content
View StLeoX's full-sized avatar

Block or report StLeoX

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Lightweight localhost observability monitor for OpenClaw: gateway status, event feed, and snapshots.

TypeScript 10 3 Updated Mar 2, 2026

Run Slurm on Kubernetes. A Slinky project.

Go 337 95 Updated Jul 28, 2026

Open-source & free — Battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in fine-tuned ruleset (NPE, thre…

Go 15,372 1,033 Updated Jul 28, 2026

Manages Unified Access to Generative AI Services built on Envoy Gateway

Go 1,873 318 Updated Jul 28, 2026

Evaluate and Enhance Your LLM Deployments for Real-World Inference Needs

Python 1,440 200 Updated Jul 28, 2026

Offline optimization of your disaggregated Dynamo graph

Python 377 142 Updated Jul 28, 2026

A workload for deploying LLM inference services on Kubernetes

Go 267 69 Updated Jul 24, 2026

Agentic RL on Any Harness at Scale

Python 714 75 Updated Jul 15, 2026

Build and run agents you can see, understand and trust.

Python 28,350 3,267 Updated Jul 28, 2026

Sub2API 一站式开源中转服务,让 Claude、Openai 、Gemini、Grok订阅统一接入,支持拼车共享,更高效分摊成本,原生工具无缝使用。

Go 34,833 7,180 Updated Jul 28, 2026

A high-performance and light-weight router for vLLM large scale deployment

Rust 331 112 Updated Jul 24, 2026

Large Language Model (LLM) Serving Paper and Resource List

29 2 Updated Jul 16, 2026

[ACL 2026] Towards Efficient Large Language Model Serving: A Survey on System-Aware KV Cache Optimization

Python 362 20 Updated Jun 28, 2026

make an agent break up a large pr into many

Python 7 1 Updated Mar 19, 2026

Generate production-quality SVG+PNG technical diagrams from natural language. 7 styles, UML support, and AI/Agent workflow patterns.

Python 9,483 796 Updated Jul 17, 2026

Hundreds of Agents. A handful of sandbox Pods

Go 32 4 Updated Jul 24, 2026

很多镜像都在国外。比如 gcr 。国内下载很慢,需要加速。致力于提供连接全世界的稳定可靠安全的容器镜像服务。

Shell 14,724 1,551 Updated Jul 22, 2026

BlitzScale Router - Distributed LLM Inference Router (Rust)

Rust 14 1 Updated May 25, 2026

Agent Substrate: the core system

Go 893 161 Updated Jul 28, 2026

A lightweight, configurable, and real-time simulator designed to mimic the behavior of vLLM without the need for GPUs or running actual heavy models.

Go 172 112 Updated Jul 27, 2026

Repo to replay Qwen trace

Rust 31 8 Updated Jan 9, 2026

Distributed KV cache scheduling & offloading libraries

Go 165 153 Updated Jul 28, 2026
Python 3 Updated Jan 26, 2026

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

C++ 6,062 1,027 Updated Jul 28, 2026

An autonomous agent that conducts deep research on any data using any LLM providers

Python 28,696 3,874 Updated Jul 18, 2026

The official implementations of Agent-as-a-Router: Agentic Model Routing for Coding Tasks.

TypeScript 1,007 17 Updated Jun 29, 2026

Fast Tokens

Rust 128 18 Updated Jun 23, 2026

Memory library for building stateful agents

Python 6,261 771 Updated Jul 27, 2026
Next