-
Abacus AI
- Vancouver, Canada
- https://aminya.github.io/
- in/amin-yahyaabadi
- All languages
- ASL
- Arduino
- Assembly
- Astro
- AutoHotkey
- BitBake
- C
- C#
- C++
- CMake
- CSS
- Clojure
- CodeQL
- CoffeeScript
- Crystal
- Cuda
- D
- Dart
- Dockerfile
- Emacs Lisp
- F#
- Fortran
- GAMS
- GDScript
- Gherkin
- Go
- HTML
- Handlebars
- Haskell
- Java
- JavaScript
- Julia
- Jupyter Notebook
- Kotlin
- LLVM
- Less
- LilyPond
- Lua
- MATLAB
- Makefile
- Markdown
- Marko
- Meson
- Mojo
- OCaml
- PHP
- Perl
- PowerShell
- Python
- QML
- Roff
- Ruby
- Rust
- SCSS
- Scala
- Scheme
- Shell
- Standard ML
- Starlark
- Svelte
- Swift
- Tcl
- TeX
- TypeScript
- VHDL
- Visual Basic .NET
- WebAssembly
- XSLT
- Zig
Starred repositories
A curated suite of AI agent skills for systems and low-level programming with C/C++, Rust, and Zig toolchains, covering compilers, debuggers, profilers, build systems, sanitizers, and binary analysis
Open source alternative to Semrush and Ahrefs
A collection of agent skills for CAD, robotics and hardware design
victortassinari / FossFLOW
Forked from leonj1/OpenFLOWMake beautiful isometric infrastructure diagrams
A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations
Clone a voice from a few minutes of audio and generate speech locally — Qwen3-TTS fine-tuning pipeline with CLI and web UI
Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.
SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.
PrismML-Eng / llama.cpp
Forked from ggml-org/llama.cppLLM inference in C/C++
SSD-streaming inference engine for giant MoE models (Rust + CUDA). GLM 5.2 743B at 2 tok/s and Hy3 295B at 7 tok/s on two consumer 16GB GPUs. Zero-config multi-GPU: measures PCIe bandwidth, places …
A library for writing reactive single page web apps
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downs…
Expose VSCode LSP capabilities to AI agents through MCP. Support multiple instances. 通过 MCP 给 AI 提供 VSCode LSP 能力,支持多实例
Windows Batch support for Sublime's LSP plugin provided through RechInformatica/rech-editor-batch.
Flyline: a Bash plugin to replace readline for a modern line editing experience: syntax highlighting, agent integration, rich prompts, tooltips, fuzzy history search, and more!
cut Fable 5 token usage by rendering text context as images
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
ZeroFS: A log-structured filesystem for S3. ZeroFS serves S3-compatible buckets as POSIX filesystems over NFS and 9P, or as raw block devices over NBD.
Fast, collaborative live terminal sharing over the web
Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
The "Missing GitHub Status Page" -- a Flat Data attempt at historically documenting GitHub statuses
Measuring frontier coding agents on original, long-horizon engineering tasks
TokenSpeed is a speed-of-light LLM inference engine.
How much experts do we need to serve a model?
Fast, lossless LLM inference via dual-view diffusion decoding.
Anbeeld / beellama.cpp
Forked from ggml-org/llama.cppKVarN, KV cache precision tail, low-bit quants in llama.cpp for longer context of better precision in the same VRAM
Monoscope lets you ingest and explore your logs, traces and metrics. We store these in S3 compatible buckets. Query in natural language via LLMs.