Skip to content
View aminya's full-sized avatar

Sponsors

@keygen-sh

Block or report aminya

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Open-source vacuum robot cleaner

Python 6,203 262 Updated Jul 25, 2026

A curated suite of AI agent skills for systems and low-level programming with C/C++, Rust, and Zig toolchains, covering compilers, debuggers, profilers, build systems, sanitizers, and binary analysis

JavaScript 147 19 Updated Jun 27, 2026

Open source alternative to Semrush and Ahrefs

TypeScript 7,817 852 Updated Jul 23, 2026

A collection of agent skills for CAD, robotics and hardware design

JavaScript 10,351 1,129 Updated Jul 11, 2026

Make beautiful isometric infrastructure diagrams

TypeScript 225 38 Updated Aug 21, 2025

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

Python 19,015 1,483 Updated Jul 24, 2026

Clone a voice from a few minutes of audio and generate speech locally — Qwen3-TTS fine-tuning pipeline with CLI and web UI

Python 142 19 Updated Jul 18, 2026

Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.

Swift 13,509 1,400 Updated Jul 24, 2026

SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.

Rust 22,450 4,240 Updated Jul 24, 2026

LLM inference in C/C++

C++ 403 76 Updated Jul 22, 2026

SSD-streaming inference engine for giant MoE models (Rust + CUDA). GLM 5.2 743B at 2 tok/s and Hy3 295B at 7 tok/s on two consumer 16GB GPUs. Zero-config multi-GPU: measures PCIe bandwidth, places …

Rust 62 8 Updated Jul 25, 2026

A flight-compliant WebAssembly interpreter.

Rust 1,399 44 Updated Jul 24, 2026

A library for writing reactive single page web apps

Rust 288 8 Updated Jul 15, 2026

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downs…

Python 3,306 511 Updated Jul 25, 2026

Expose VSCode LSP capabilities to AI agents through MCP. Support multiple instances. 通过 MCP 给 AI 提供 VSCode LSP 能力,支持多实例

TypeScript 37 10 Updated Jul 25, 2026

Windows Batch support for Sublime's LSP plugin provided through RechInformatica/rech-editor-batch.

Shell 6 Updated Jun 27, 2026

Flyline: a Bash plugin to replace readline for a modern line editing experience: syntax highlighting, agent integration, rich prompts, tooltips, fuzzy history search, and more!

Rust 962 21 Updated Jul 20, 2026

cut Fable 5 token usage by rendering text context as images

TypeScript 6,692 571 Updated Jul 25, 2026

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

Python 62,312 4,705 Updated Jul 25, 2026

ZeroFS: A log-structured filesystem for S3. ZeroFS serves S3-compatible buckets as POSIX filesystems over NFS and 9P, or as raw block devices over NBD.

Rust 2,905 105 Updated Jul 24, 2026

Fast, collaborative live terminal sharing over the web

Rust 7,558 296 Updated Jun 19, 2025

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

JavaScript 89,207 4,891 Updated Jul 15, 2026

The "Missing GitHub Status Page" -- a Flat Data attempt at historically documenting GitHub statuses

HTML 501 21 Updated Jul 25, 2026

Measuring frontier coding agents on original, long-horizon engineering tasks

Python 1,228 76 Updated Jul 22, 2026
Python 305 7 Updated Jul 2, 2026

TokenSpeed is a speed-of-light LLM inference engine.

Python 1,658 199 Updated Jul 25, 2026

How much experts do we need to serve a model?

Python 152 15 Updated Mar 18, 2026

Fast, lossless LLM inference via dual-view diffusion decoding.

Python 460 20 Updated May 18, 2026

KVarN, KV cache precision tail, low-bit quants in llama.cpp for longer context of better precision in the same VRAM

C++ 803 41 Updated Jul 25, 2026

Monoscope lets you ingest and explore your logs, traces and metrics. We store these in S3 compatible buckets. Query in natural language via LLMs.

Haskell 1,441 62 Updated Jul 25, 2026
Next