Skip to content
View ritazh's full-sized avatar
🎯
Focusing
🎯
Focusing

Organizations

@kubernetes @open-policy-agent @virtual-kubelet @kubernetes-sigs @coreweave

Block or report ritazh

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

slime is an LLM post-training framework for RL Scaling.

Python 7,754 1,116 Updated Aug 4, 2026

An LLM post-training framework with vLLM for RL Scaling

Python 400 73 Updated Aug 3, 2026

Transaction Tokens (RFC 8693) for Kubernetes — seal identity, context, and authorization across multi-hop agent workflows via AgentGateway + ext_authz.

Go 25 7 Updated Jun 12, 2026

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downs…

Python 3,380 526 Updated Aug 4, 2026

Agent skill that removes signs of AI-generated writing from text

Python 33,179 3,013 Updated Jul 22, 2026

The best-benchmarked open-source AI memory system. And it's free.

Python 58,040 7,458 Updated Aug 3, 2026

llm-d benchmark scripts and tooling

Python 64 119 Updated Aug 4, 2026

Tooling for optimized, validated, and reproducible GPU-accelerated AI runtime in Kubernetes

Go 364 76 Updated Aug 4, 2026

A cloud-agnostic Kubernetes node autoscaler that dynamically scales infrastructure across Azure and emerging neoclouds like Nebius—managed from a single control plane.

Go 9 2 Updated Jun 24, 2026

Rally your AI squad to GitHub issues and PRs via git worktrees

JavaScript 34 2 Updated Jul 31, 2026
Shell 2 1 Updated Jul 25, 2026

💫 Toolkit to help you get started with Spec-Driven Development

Python 125,239 11,192 Updated Aug 3, 2026

Inspektor Gadget is a set of tools and framework for data collection and system inspection on Kubernetes clusters and Linux hosts using eBPF

C 2,903 364 Updated Aug 4, 2026

✈️ Kubernetes-native platform for deploying and managing AI inference across multiple providers

TypeScript 95 33 Updated Jul 28, 2026

Discover ingress-nginx usage and auto-generate Gateway API migration plans before ingress-nginx reaches end-of-life (March 2026).

Go 16 Updated Nov 26, 2025

The AI Compute Platform for frontier teams. SkyPilot turns fragmented AI compute into one AI supercomputer, so frontier AI teams build custom intelligence faster.

Python 10,442 1,175 Updated Aug 4, 2026

The best ChatGPT that $100 can buy.

Python 56,950 7,884 Updated Aug 2, 2026

A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)

Python 3,640 272 Updated Feb 8, 2026

A sample pack of GitHub Agentic Workflows!

Makefile 881 127 Updated Jul 29, 2026
2 Updated Sep 26, 2025

Achieve state of the art inference performance with modern accelerators on Kubernetes

Shell 3,967 649 Updated Aug 4, 2026

Wassette: A security-oriented runtime that runs WebAssembly Components via MCP

Rust 931 69 Updated Aug 4, 2026

gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI

Python 20,283 2,135 Updated Jul 24, 2026

Next Generation Agentic Proxy for AI Agents and MCP servers

Rust 4,209 702 Updated Aug 3, 2026

LLM inference in C/C++

C++ 122,654 21,307 Updated Aug 4, 2026

Home of the out-of-tree KAITO plugin for Headlamp Kubernetes UI

TypeScript 7 5 Updated Aug 8, 2025

The Security Toolkit for LLM Interactions

Python 3,202 434 Updated Jul 8, 2026

Set of tools to assess and improve LLM security.

Python 4,328 766 Updated Jul 27, 2026

A comprehensive social media management tool designed to help you create, format, and post content across multiple platforms including LinkedIn, Twitter/X, Bluesky, and Mastodon. Features advanced …

TypeScript 100 15 Updated Jan 15, 2026

Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton

Go 485 89 Updated Aug 3, 2026
Next