Skip to content
View ritazh's full-sized avatar
🎯
Focusing
🎯
Focusing

Organizations

@kubernetes @open-policy-agent @virtual-kubelet @kubernetes-sigs @coreweave

Block or report ritazh

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

slime is an LLM post-training framework for RL Scaling.

Python 7,832 1,130 Updated Aug 10, 2026

An LLM post-training framework with vLLM for RL Scaling

Python 408 74 Updated Aug 3, 2026

Transaction Tokens (RFC 8693) for Kubernetes — seal identity, context, and authorization across multi-hop agent workflows via AgentGateway + ext_authz.

Go 25 8 Updated Jun 12, 2026

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downs…

Python 3,418 538 Updated Aug 10, 2026

Agent skill that removes signs of AI-generated writing from text

Python 34,688 3,118 Updated Jul 22, 2026

The best-benchmarked open-source AI memory system. And it's free.

Python 58,279 7,494 Updated Aug 8, 2026

llm-d benchmark scripts and tooling

Python 64 123 Updated Aug 10, 2026

Tooling for optimized, validated, and reproducible GPU-accelerated AI runtime in Kubernetes

Go 380 79 Updated Aug 10, 2026

A cloud-agnostic Kubernetes node autoscaler that dynamically scales infrastructure across Azure and emerging neoclouds like Nebius—managed from a single control plane.

Go 9 2 Updated Jun 24, 2026

Rally your AI squad to GitHub issues and PRs via git worktrees

JavaScript 34 2 Updated Aug 6, 2026
Shell 2 1 Updated Jul 25, 2026

💫 Toolkit to help you get started with Spec-Driven Development

Python 126,084 11,267 Updated Aug 10, 2026

Inspektor Gadget is a set of tools and framework for data collection and system inspection on Kubernetes clusters and Linux hosts using eBPF

C 2,906 365 Updated Aug 10, 2026

✈️ Kubernetes-native platform for deploying and managing AI inference across multiple providers

TypeScript 97 33 Updated Aug 7, 2026

Discover ingress-nginx usage and auto-generate Gateway API migration plans before ingress-nginx reaches end-of-life (March 2026).

Go 16 Updated Nov 26, 2025

The AI Compute Platform for frontier teams. SkyPilot turns fragmented AI compute into one AI supercomputer, so frontier AI teams build custom intelligence faster.

Python 10,465 1,179 Updated Aug 10, 2026

The best ChatGPT that $100 can buy.

Python 57,106 7,912 Updated Aug 2, 2026

A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)

Python 3,657 273 Updated Feb 8, 2026

A sample pack of GitHub Agentic Workflows!

Makefile 893 130 Updated Jul 29, 2026
2 Updated Sep 26, 2025

Achieve state of the art inference performance with modern accelerators on Kubernetes

Shell 4,005 666 Updated Aug 10, 2026

Wassette: A security-oriented runtime that runs WebAssembly Components via MCP

Rust 934 70 Updated Aug 5, 2026

gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI

Python 20,300 2,131 Updated Jul 24, 2026

Next Generation Agentic Proxy for AI Agents and MCP servers

Rust 4,292 716 Updated Aug 10, 2026

LLM inference in C/C++

C++ 123,327 21,519 Updated Aug 10, 2026

Home of the out-of-tree KAITO plugin for Headlamp Kubernetes UI

TypeScript 7 5 Updated Aug 8, 2025

The Security Toolkit for LLM Interactions

Python 3,201 436 Updated Jul 8, 2026

Set of tools to assess and improve LLM security.

Python 4,339 767 Updated Aug 6, 2026

A comprehensive social media management tool designed to help you create, format, and post content across multiple platforms including LinkedIn, Twitter/X, Bluesky, and Mastodon. Features advanced …

TypeScript 100 15 Updated Jan 15, 2026

Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton

Go 487 89 Updated Aug 10, 2026
Next