Skip to content
View aminya's full-sized avatar

Sponsors

@keygen-sh

Block or report aminya

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Open-source vacuum robot cleaner

Python 7,062 331 Updated Aug 3, 2026

A curated suite of AI agent skills for systems and low-level programming with C/C++, Rust, and Zig toolchains, covering compilers, debuggers, profilers, build systems, sanitizers, and binary analysis

JavaScript 155 20 Updated Jun 27, 2026

Open source alternative to Semrush and Ahrefs

TypeScript 10,327 1,191 Updated Jul 30, 2026

A library of agent skills for CAD, CAE and CAM

JavaScript 12,781 1,345 Updated Aug 3, 2026

Make beautiful isometric infrastructure diagrams

TypeScript 250 42 Updated Aug 21, 2025

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

Python 19,167 1,507 Updated Aug 2, 2026

Clone a voice from a few minutes of audio and generate speech locally — Qwen3-TTS fine-tuning pipeline with CLI and web UI

Python 147 21 Updated Jul 18, 2026

Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.

Swift 13,589 1,455 Updated Jul 24, 2026

SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.

Rust 24,083 4,564 Updated Aug 3, 2026

LLM inference in C/C++

C++ 441 82 Updated Aug 3, 2026

SSD-streaming inference engine for giant MoE models (Rust + CUDA). GLM 5.2 743B at 2 tok/s and Hy3 295B at 7 tok/s on two consumer 16GB GPUs. Zero-config multi-GPU: measures PCIe bandwidth, places …

Rust 163 19 Updated Aug 3, 2026

A flight-compliant WebAssembly interpreter.

Rust 1,458 48 Updated Aug 1, 2026

A library for writing reactive single page web apps

Rust 289 8 Updated Jul 26, 2026

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downs…

Python 3,381 527 Updated Aug 4, 2026

Expose VS Code language intelligence to AI agents via MCP, with automatic routing across multiple windows

TypeScript 39 11 Updated Jul 27, 2026

Windows Batch support for Sublime's LSP plugin provided through RechInformatica/rech-editor-batch.

Shell 6 Updated Jun 27, 2026

Flyline: a Bash plugin to replace readline for a modern line editing experience: syntax highlighting, agent integration, rich prompts, tooltips, fuzzy history search, and more!

Rust 1,015 27 Updated Aug 3, 2026

cut Claude Code token usage by rendering text context as images

TypeScript 6,944 599 Updated Aug 4, 2026

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

Python 64,566 4,914 Updated Aug 4, 2026

ZeroFS: A log-structured filesystem for S3. ZeroFS serves S3-compatible buckets as POSIX filesystems over NFS and 9P, or as raw block devices over NBD.

Rust 2,937 108 Updated Aug 3, 2026

Fast, collaborative live terminal sharing over the web

Rust 7,609 300 Updated Jun 19, 2025

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

JavaScript 95,520 5,251 Updated Jul 15, 2026

The "Missing GitHub Status Page" -- a Flat Data attempt at historically documenting GitHub statuses

HTML 504 21 Updated Aug 4, 2026

Measuring frontier coding agents on original, long-horizon engineering tasks

Python 1,311 84 Updated Jul 22, 2026
Python 305 7 Updated Jul 2, 2026

TokenSpeed is a speed-of-light LLM inference engine.

Python 1,802 216 Updated Aug 4, 2026

How much experts do we need to serve a model?

Python 152 15 Updated Mar 18, 2026

Fast, lossless LLM inference via dual-view diffusion decoding.

Python 466 20 Updated May 18, 2026

KVarN, KV cache precision tail, low-bit quants in llama.cpp for longer context of better precision in the same VRAM

C++ 830 41 Updated Aug 2, 2026

Monoscope lets you ingest and explore your logs, traces and metrics. We store these in S3 compatible buckets. Query in natural language via LLMs.

Haskell 1,499 64 Updated Aug 4, 2026
Next