Skip to content
View Gforky's full-sized avatar

Block or report Gforky

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

DeepSeek Harness: Everything is a Plugin.

TypeScript 84,720 7,484 Updated Aug 13, 2026

Browser automation CLI for AI agents

Rust 40,611 2,676 Updated Aug 13, 2026

Production-ready MoE load balancing via real-time expert replication

Cuda 238 11 Updated Jul 17, 2026

Litmus helps SREs and developers practice chaos engineering in a Cloud-native way. Chaos experiments are published at the ChaosHub (https://hub.litmuschaos.io). Community notes is at https://hackmd…

Go 5,594 890 Updated Jul 31, 2026

KernelBench: Can LLMs Write GPU Kernels? - Benchmark + Toolkit with Torch -> CUDA (+ more DSLs)

Jupyter Notebook 1,199 188 Updated Mar 24, 2026

Agent-assisted and full-agent reproducibility package for MLSys 2026 FlashInfer AI Kernel Generation Contest submissions: kernels, agent workflows, skills, configs, writeup, benchmark artifacts, an…

Python 21 2 Updated Jul 21, 2026

Mixture-of-experts (MoE) training megakernel for NVL72s

Python 530 60 Updated Aug 14, 2026

Expert Parallelism Load Balancer

Python 1,419 206 Updated Mar 24, 2025

maximal update parametrization (µP)

Jupyter Notebook 1,749 105 Updated Jul 17, 2024

[NSDI25] AutoCCL: Automated Collective Communication Tuning for Accelerating Distributed and Parallel DNN Training

C++ 35 3 Updated May 2, 2025

Communication patterns for AI, built on top of NCCL device and host APIs

Cuda 28 8 Updated Aug 12, 2026

SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.

Rust 25,143 4,757 Updated Aug 13, 2026

A CPU+GPU Profiling library that provides access to timeline traces and hardware performance counters.

C++ 987 267 Updated Aug 13, 2026

A collection of tricks and tools to speed up transformer models

TeX 220 15 Updated Aug 11, 2026

Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

C 24,626 2,688 Updated Aug 14, 2026
Python 907 86 Updated Aug 13, 2026

Skill to give Claude Code (and any coding agent) the ability to generate beautiful and practical Excalidraw diagrams.

Python 4,418 502 Updated Mar 1, 2026

Generate beautiful dark-themed system architecture diagrams as standalone HTML/SVG files. Works as a Claude AI skill.

HTML 6,911 533 Updated May 13, 2026

Venus Collective Communication Library, supported by SII and Infrawaves.

C++ 150 8 Updated Jun 24, 2026

Agent skill for beautiful, verifiable architecture, workflow, sequence, data-flow, and lifecycle diagrams—self-contained HTML with motion and crisp export.

HTML 12,275 925 Updated Aug 14, 2026

Generate draw.io diagrams from natural language — 11 presets (UML, SysML/MBSE, BPMN, network, C4…), 36 tools: codebase/CI/infra-to-diagram, image→editable diagram, mind maps, build-up animation, ex…

Python 7,619 552 Updated Aug 5, 2026

Open Machine Learning Compiler Framework

Python 13,664 3,953 Updated Aug 12, 2026

Agent for collecting, processing, aggregating, and writing metrics, logs, and other arbitrary data.

Go 17,750 5,834 Updated Aug 13, 2026

Kubernetes Operator for OpenTelemetry Collector

Go 1,747 647 Updated Aug 14, 2026

A high-performance distributed deep learning system targeting large-scale and automated distributed training.

Python 340 43 Updated Dec 13, 2025

ByteCheckpoint: An Unified Checkpointing Library for LFMs

Python 289 21 Updated Feb 2, 2026

Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.

Python 1,259 97 Updated Aug 13, 2026

🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman

Go 98,106 5,676 Updated Aug 13, 2026

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

JavaScript 102,486 5,642 Updated Aug 7, 2026
Next