Skip to content
View Gforky's full-sized avatar

Block or report Gforky

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

DeepSeek Harness: Everything is a Plugin.

TypeScript 156,895 16,294 Updated Aug 17, 2026

Browser automation CLI for AI agents

Rust 40,858 2,710 Updated Aug 17, 2026

Production-ready MoE load balancing via real-time expert replication

Cuda 240 11 Updated Jul 17, 2026

Litmus helps SREs and developers practice chaos engineering in a Cloud-native way. Chaos experiments are published at the ChaosHub (https://hub.litmuschaos.io). Community notes is at https://hackmd…

Go 5,597 889 Updated Jul 31, 2026

KernelBench: Can LLMs Write GPU Kernels? - Benchmark + Toolkit with Torch -> CUDA (+ more DSLs)

Jupyter Notebook 1,205 189 Updated Mar 24, 2026

Agent-assisted and full-agent reproducibility package for MLSys 2026 FlashInfer AI Kernel Generation Contest submissions: kernels, agent workflows, skills, configs, writeup, benchmark artifacts, an…

Python 21 2 Updated Jul 21, 2026

Mixture-of-experts (MoE) training megakernel for NVL72s

Python 544 62 Updated Aug 14, 2026

Expert Parallelism Load Balancer

Python 1,419 208 Updated Mar 24, 2025

maximal update parametrization (µP)

Jupyter Notebook 1,751 105 Updated Jul 17, 2024

[NSDI25] AutoCCL: Automated Collective Communication Tuning for Accelerating Distributed and Parallel DNN Training

C++ 35 3 Updated May 2, 2025

Communication patterns for AI, built on top of NCCL device and host APIs

Cuda 30 8 Updated Aug 16, 2026

SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.

Rust 25,584 4,817 Updated Aug 17, 2026

A CPU+GPU Profiling library that provides access to timeline traces and hardware performance counters.

C++ 988 267 Updated Aug 14, 2026

A collection of tricks and tools to speed up transformer models

TeX 221 15 Updated Aug 11, 2026

Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

C 25,380 2,749 Updated Aug 18, 2026
Python 916 87 Updated Aug 16, 2026

Skill to give Claude Code (and any coding agent) the ability to generate beautiful and practical Excalidraw diagrams.

Python 4,464 507 Updated Mar 1, 2026

Generate beautiful dark-themed system architecture diagrams as standalone HTML/SVG files. Works as a Claude AI skill.

HTML 6,953 531 Updated May 13, 2026

Venus Collective Communication Library, supported by SII and Infrawaves.

C++ 150 8 Updated Jun 24, 2026

Agent skill for beautiful, verifiable architecture, workflow, sequence, data-flow, and lifecycle diagrams—self-contained HTML with motion and crisp export.

HTML 14,091 1,032 Updated Aug 17, 2026

Generate draw.io diagrams from natural language — 11 presets (UML, SysML/MBSE, BPMN, network, C4…), 36 tools: codebase/CI/infra-to-diagram, image→editable diagram, mind maps, build-up animation, ex…

Python 7,761 557 Updated Aug 5, 2026

Open Machine Learning Compiler Framework

Python 13,669 3,957 Updated Aug 18, 2026

Agent for collecting, processing, aggregating, and writing metrics, logs, and other arbitrary data.

Go 17,756 5,836 Updated Aug 18, 2026

Kubernetes Operator for OpenTelemetry Collector

Go 1,746 648 Updated Aug 18, 2026

A high-performance distributed deep learning system targeting large-scale and automated distributed training.

Python 340 43 Updated Dec 13, 2025

ByteCheckpoint: An Unified Checkpointing Library for LFMs

Python 289 21 Updated Feb 2, 2026

Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.

Python 1,259 99 Updated Aug 13, 2026

🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman

Go 98,847 5,725 Updated Aug 18, 2026

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

JavaScript 105,058 5,798 Updated Aug 7, 2026
Next