Lists (7)
Sort Name ascending (A-Z)
- All languages
- Assembly
- Astro
- C
- C#
- C++
- CSS
- Clojure
- ColdFusion
- Cuda
- Cython
- Dart
- Dockerfile
- Elixir
- Erlang
- Go
- HCL
- HTML
- Java
- JavaScript
- Julia
- Jupyter Notebook
- Kotlin
- LLVM
- Lean
- LookML
- Lua
- MATLAB
- MLIR
- Makefile
- Markdown
- Metal
- Mojo
- Nix
- OCaml
- Objective-C
- PHP
- PLpgSQL
- Processing
- Python
- R
- Ruby
- Rust
- SCSS
- Scala
- Shell
- Stan
- Swift
- SystemVerilog
- TeX
- TypeScript
- Vue
- Zig
Starred repositories
Mixture-of-experts (MoE) training megakernel for NVL72s
Pure Rust multimedia format demuxing, tag reading, and audio decoding library
An asynchronous runtime for writing applications and services. Supports io_uring, epoll, kqueue, and poll for I/O.
Fast Rust library for PDF inspection, classification, and text extraction. Intelligently detects scanned vs text-based PDFs to enable smart routing decisions.
📰 Must-read papers on KV Cache Compression (constantly updating 🤗).
Static @fit checks for ordinary TypeScript layout code
Lightweight coding agent that runs in your terminal
Train speculative decoding models effortlessly and port them smoothly to SGLang serving.
High-performance GPU kernels written in TIRx.
A tutorial on modern GPU programming for machine learning systems
DFlash: Block Diffusion for Flash Speculative Decoding
Official PyTorch Implementation for Readout Guidance, CVPR 2024
An executable specification language with delightful tooling based on the temporal logic of actions (TLA)
A digital camera you can build yourself with Codex.
A lightweight, lightning-fast, in-process vector database
Learn it. Build it. Ship it for others.
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
Helpful kernel tutorials, examples and SKILLs for tile-based GPU programming
Official repository for our work on micro-budget training of large-scale diffusion models.
The agent that grows with you
AI-native design editor. Open-source Figma alternative.
Zig INferenCe Engine — Local LLM inference on AMD GPUs and Apple Silicon
[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
An MLIR-based compiler that takes GPU kernels and compiles them to real hardware instructions. Interactive web visualizer included.
Course on Flash-attention in Triton