Skip to content
View zhyncs's full-sized avatar
🎯
🎯

Sponsoring

@lightseekorg

Organizations

@lightseekorg @smg-project

Block or report zhyncs

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI

TypeScript 85,830 10,662 Updated Aug 9, 2026

AgentENV (AENV) is a distributed platform for running agent environments at scale.

Rust 2,980 240 Updated Aug 8, 2026

MoonEP: A Perfectly Balanced Expert Parallelism Library via Dynamic Redundant Experts

Python 1,053 111 Updated Aug 7, 2026

Inspect: A framework for large language model evaluations

Python 2,510 637 Updated Aug 9, 2026

TokenSpeed is a speed-of-light LLM inference engine.

Python 1,836 225 Updated Aug 9, 2026

A PyTorch native library for training speculative decoding models

Python 224 59 Updated Aug 9, 2026

Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing,…

Rust 452 133 Updated Aug 9, 2026

A program which ensures source code files have copyright license headers by scanning directory patterns recursively

Go 884 195 Updated Oct 28, 2025

Open Visual Agentic Intelligence

2,297 323 Updated Aug 6, 2026

Lightweight coding agent that runs in your terminal

Rust 104,888 15,871 Updated Aug 9, 2026

cocoon

C++ 874 124 Updated May 21, 2026

🐶 Kubernetes CLI To Manage Your Clusters In Style!

Go 34,304 2,244 Updated Aug 8, 2026

The easiest, most secure way to use WireGuard and 2FA.

Go 34,966 3,047 Updated Aug 8, 2026

DeepGEMM: clean and efficient BLAS kernel library on GPU

Cuda 7,636 1,159 Updated Jul 20, 2026

DeepEP: an efficient expert-parallel communication library

Cuda 9,965 1,368 Updated Aug 5, 2026

FlashMLA: Efficient Multi-head Latent Attention Kernels

C++ 12,829 1,119 Updated Jul 28, 2026

Fast, Flexible and Portable Structured Generation

C++ 1,813 182 Updated Aug 6, 2026

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

Python 14,342 2,649 Updated Aug 9, 2026

Official Implementation of EAGLE-1 (ICML'24), EAGLE-2 (EMNLP'24), and EAGLE-3 (NeurIPS'25).

Python 2,500 296 Updated Feb 20, 2026

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

C++ 6,208 1,066 Updated Aug 9, 2026

High-Performance C++ Fundamental Library

C++ 651 101 Updated Mar 16, 2026

Development repository for the Triton language and compiler

MLIR 19,905 3,087 Updated Aug 9, 2026

FlashInfer: Kernel Library for LLM Serving

Python 6,133 1,250 Updated Aug 9, 2026

CUDA Templates and Python DSLs for High-Performance Linear Algebra

C++ 10,215 2,001 Updated Aug 8, 2026

Fast and memory-efficient exact attention

Python 24,656 2,972 Updated Aug 9, 2026

brpc is an Industrial-grade RPC framework using C++ Language, which is often used in high performance system such as Search, Storage, Machine learning, Advertisement, Recommendation etc. "brpc" mea…

C++ 17,576 4,129 Updated Aug 9, 2026