Skip to content
View james0zan's full-sized avatar

Organizations

@kvcache-ai

Block or report james0zan

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

AgentENV (AENV) is a distributed platform for running agent environments at scale.

Rust 10 6 Updated Jul 27, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 30,789 7,427 Updated Jul 27, 2026

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

Python 19,048 1,485 Updated Jul 24, 2026

A PyTorch native library for training speculative decoding models

Python 207 52 Updated Jul 27, 2026

Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing,…

Rust 418 126 Updated Jul 26, 2026

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

Python 1,796 327 Updated Jul 27, 2026

A visualized theorem prover based on Lean 4

TypeScript 8 Updated Nov 12, 2025

Proof the completeness of Russel's Axiomatic System in lean4, and using C++ to automatically convert lean4 file to markdown file

Lean 4 Updated Jan 6, 2026
Go 95 8 Updated Sep 15, 2025

Checkpoint-engine is a simple middleware to update model weights in LLM inference engines

Python 983 100 Updated Jul 4, 2026

A Distributed Attention Towards Linear Scalability for Ultra-Long Context, Heterogeneous Data Training

Python 890 63 Updated Jul 27, 2026

LMCache: Supercharge Your LLM with the Fastest KV Cache Layer

Python 10,906 1,624 Updated Jul 27, 2026

I created a claude code deep researcher that seems to work better than the current deep research models

Python 146 26 Updated Jun 9, 2026

MoBA: Mixture of Block Attention for Long-Context LLMs

Python 2,154 157 Updated Apr 3, 2025

High-speed Large Language Model Serving for Local Deployment

C++ 9,681 588 Updated May 11, 2026

Merico Build is a web app empowering open source developers, maintainers, and communities with metrics from Git, GitHub, and more.

494 24 Updated Jun 29, 2021

CSI driver to bring SPDK to Kubernetes storage through NVMe-oF or iSCSI. Supports dynamic volume provisioning and enables Pods to use SPDK storage transparently.

Go 88 45 Updated Jan 27, 2026

A RocksDB compatible KV storage engine with better performance

C++ 2,152 213 Updated Jul 13, 2026

Concurrent data structures in C++

C++ 1,458 159 Updated May 16, 2026

Bot Framework provides the most comprehensive experience for building conversation applications.

JavaScript 7,808 2,426 Updated Dec 29, 2025

Multithreaded HTTP Download Accelerator

C 23 12 Updated Jul 27, 2014

A Python library for using the duoshuo API

Python 3 Updated Jul 22, 2012

A Python library for using the duoshuo API

Python 88 31 Updated Nov 23, 2021

PyCoder's Weekly Chinese Translate Sources Repo

HTML 393 92 Updated Dec 3, 2017