Skip to content
View aminya's full-sized avatar

Sponsors

@keygen-sh

Block or report aminya

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Open-source vacuum robot cleaner

Python 8,140 377 Updated Aug 7, 2026

A curated suite of AI agent skills for systems and low-level programming with C/C++, Rust, and Zig toolchains, covering compilers, debuggers, profilers, build systems, sanitizers, and binary analysis

JavaScript 157 20 Updated Jun 27, 2026

Open source alternative to Semrush and Ahrefs

TypeScript 10,821 1,243 Updated Jul 30, 2026

A library of agent skills for CAD, CAE and CAM

JavaScript 13,026 1,376 Updated Aug 6, 2026

Make beautiful isometric infrastructure diagrams

TypeScript 255 43 Updated Aug 21, 2025

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

Python 19,199 1,510 Updated Aug 6, 2026

Clone a voice from a few minutes of audio and generate speech locally — Qwen3-TTS fine-tuning pipeline with CLI and web UI

Python 147 21 Updated Jul 18, 2026

Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.

Swift 13,626 1,467 Updated Jul 24, 2026

SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.

Rust 24,373 4,629 Updated Aug 6, 2026

LLM inference in C/C++

C++ 447 86 Updated Aug 6, 2026

SSD-streaming inference engine for giant MoE models (Rust + CUDA). GLM 5.2 743B at 2 tok/s and Hy3 295B at 7 tok/s on two consumer 16GB GPUs. Zero-config multi-GPU: measures PCIe bandwidth, places …

Rust 193 24 Updated Aug 7, 2026

A flight-compliant WebAssembly interpreter.

Rust 1,461 49 Updated Aug 7, 2026

A library for writing reactive single page web apps

Rust 289 8 Updated Jul 26, 2026

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downs…

Python 3,405 536 Updated Aug 7, 2026

Expose VS Code language intelligence to AI agents via MCP, with automatic routing across multiple windows

TypeScript 39 11 Updated Jul 27, 2026

Windows Batch support for Sublime's LSP plugin provided through RechInformatica/rech-editor-batch.

Shell 6 Updated Jun 27, 2026

Flyline: a Bash plugin to replace readline for a modern line editing experience: syntax highlighting, agent integration, rich prompts, tooltips, fuzzy history search, and more!

Rust 1,024 27 Updated Aug 6, 2026

cut Claude Code token usage by rendering text context as images

TypeScript 6,981 604 Updated Aug 6, 2026

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

Python 65,361 4,982 Updated Aug 7, 2026

ZeroFS: A log-structured filesystem for S3. ZeroFS serves S3-compatible buckets as POSIX filesystems over NFS and 9P, or as raw block devices over NBD.

Rust 2,943 108 Updated Aug 7, 2026

Fast, collaborative live terminal sharing over the web

Rust 7,631 300 Updated Jun 19, 2025

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

JavaScript 98,103 5,389 Updated Jul 15, 2026

The "Missing GitHub Status Page" -- a Flat Data attempt at historically documenting GitHub statuses

JavaScript 508 21 Updated Aug 7, 2026

Measuring frontier coding agents on original, long-horizon engineering tasks

Python 1,333 84 Updated Aug 6, 2026
Python 305 7 Updated Jul 2, 2026

TokenSpeed is a speed-of-light LLM inference engine.

Python 1,825 221 Updated Aug 7, 2026

How much experts do we need to serve a model?

Python 152 15 Updated Mar 18, 2026

Fast, lossless LLM inference via dual-view diffusion decoding.

Python 471 21 Updated May 18, 2026

KVarN, KV cache precision tail, low-bit quants in llama.cpp for longer context of better precision in the same VRAM

C++ 838 41 Updated Aug 6, 2026

Monoscope lets you ingest and explore your logs, traces and metrics. We store these in S3 compatible buckets. Query in natural language via LLMs.

Haskell 1,514 65 Updated Aug 7, 2026
Next