Skip to content
View bretagne-peiqi's full-sized avatar
  • Shanghai, China

Block or report bretagne-peiqi

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

trendigger.com is a trend-tracking platform for aggregating and visualizing what people are searching on Google. It provides a hour-by-hour breakdown of popular keywords across different countries …

JavaScript 4 Updated Sep 6, 2026

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

Python 19,536 1,575 Updated Sep 23, 2026

Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing,…

Rust 544 172 Updated Sep 24, 2026

🤖 AI Gateway | AI Native API Gateway

Go 9,451 1,315 Updated Sep 24, 2026

A high-performance distributed file system designed to address the challenges of AI training and inference workloads.

C++ 10,225 1,098 Updated May 7, 2026

Superfast AI decision making and intelligent processing of multi-modal data.

Python 3,922 373 Updated Sep 12, 2026

Open-source, secure environment with real-world tools for enterprise-grade agents.

Python 13,943 1,047 Updated Sep 22, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 122 11 Updated Sep 24, 2026

The open source coding agent.

TypeScript 209,780 27,697 Updated Sep 24, 2026

InternRobotics' open platform for building generalized navigation foundation models.

Jupyter Notebook 1,121 143 Updated Mar 10, 2026

DeepGEMM: clean and efficient BLAS kernel library on GPU

Cuda 7,865 1,275 Updated Sep 24, 2026

Multi-Joint dynamics with Contact. A general purpose physics simulator.

C++ 15,318 1,782 Updated Sep 24, 2026

Fast and memory-efficient exact attention

Python 25,005 3,098 Updated Sep 24, 2026

Fast and memory-efficient exact attention

Python 23 35 Updated Jun 26, 2026

Distributed AI Model Training and LLM Fine-Tuning on Kubernetes

Go 2,228 1,055 Updated Sep 21, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 75,005 9,187 Updated Sep 14, 2026

The AI that really does things. Any OS. Any Platform. The lobster way. 🦞

TypeScript 390,374 82,126 Updated Sep 24, 2026

Optimized primitives for collective multi-GPU communication

C++ 5,116 1,426 Updated Sep 23, 2026

A CNI IPAM plugin that assigns IP addresses cluster-wide

Go 398 165 Updated Sep 18, 2026

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

C++ 6,651 1,266 Updated Sep 24, 2026

NVIDIA Network Operator

Go 373 82 Updated Sep 24, 2026

A unified AI model hub for aggregation & distribution. It supports cross-converting various LLMs into OpenAI-compatible, Claude-compatible, or Gemini-compatible formats. A centralized gateway for p…

Go 48,843 11,721 Updated Sep 23, 2026

Lightweight, Modular, Kubernetes-native AI serving platform for scalable model serving.

Go 467 205 Updated Sep 24, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 36,391 9,113 Updated Sep 24, 2026

A Go implementation of the Model Context Protocol (MCP), enabling seamless integration between LLM applications and external data sources and tools.

Go 9,139 884 Updated Sep 23, 2026

A workload for deploying LLM inference services on Kubernetes

Go 300 81 Updated Sep 24, 2026

DeepEP: an efficient expert-parallel communication library

Cuda 10,198 1,449 Updated Sep 23, 2026

A cert-manager webhook completing DNS01 challenge by using External DNS

Go 7 3 Updated Nov 2, 2022

Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton

Go 512 98 Updated Sep 23, 2026

Automatically provision and manage TLS certificates in Kubernetes

Go 14,088 2,452 Updated Sep 24, 2026
Next