Skip to content
View achalpandeyy's full-sized avatar
🚀
Focusing
🚀
Focusing

Block or report achalpandeyy

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

SilentPatch for GTA III, Vice City, and San Andreas

C++ 1,101 54 Updated Aug 8, 2026

Mixture-of-experts (MoE) training megakernel for NVL72s

Python 520 57 Updated Aug 10, 2026

MSLK (Meta Superintelligence Labs Kernels) is a collection of PyTorch GPU operator libraries that are designed and optimized for GenAI training and inference, such as FP8 row-wise quantization and …

Python 145 74 Updated Aug 12, 2026

Arm Performix: https://developer.arm.com/servers-and-cloud-computing/arm-performix

Go 15 3 Updated Aug 4, 2026

Reference implementation and examples of the CuTe Layout representation and algebra.

Python 264 25 Updated Aug 6, 2026

MoonEP: A Perfectly Balanced Expert Parallelism Library via Dynamic Redundant Experts

Python 1,067 118 Updated Aug 7, 2026

Flash-Flash KDA: H100-optimized Flash Kimi Delta Attention kernels

Cuda 16 Updated Jul 24, 2026

Language model tokenization at GB/s

Rust 3,963 205 Updated Aug 6, 2026

A feed-forward 3D foundation model for reconstructing scenes from streaming data

Python 16,430 1,834 Updated Aug 12, 2026

SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.

Rust 24,718 4,704 Updated Aug 11, 2026

Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

C 24,143 2,621 Updated Aug 11, 2026

A web engine for pencil-style rendering of 3D scenes to SVG — exact silhouettes, hidden-line ghosting, hatching. Great for abstract and technical scenes.

TypeScript 224 8 Updated Jul 25, 2026

Cross-platform instrumentation and introspection library written in C

C 1,002 368 Updated Aug 11, 2026

minimal cross-platform standalone C headers

C 10,191 661 Updated Aug 11, 2026

Cross-Platform HW accelerated CRC32c and CRC32 with fallback to efficient SW implementations. C interface with language bindings for each of our SDKs

C 70 59 Updated Aug 10, 2026

Random stuff about lower level iOS

C++ 476 43 Updated Jul 23, 2026

Standalone C++/GGML runtime for ThinkSound text->sound-effect generation

C++ 36 1 Updated Jul 13, 2026

Box3D is a 3D physics engine for games

C 6,017 295 Updated Aug 12, 2026

A comprehensive collection of IQA papers

TeX 1,540 90 Updated Aug 2, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,860 1,132 Updated Aug 12, 2026

High-performance code intelligence MCP server. Indexes codebases into a persistent knowledge graph — average repo in milliseconds. 158 languages, sub-ms queries, 99% fewer tokens. Single static bin…

C 38,614 3,075 Updated Aug 12, 2026

Low-latency Rust thread pool with parallel iterators

Rust 94 4 Updated Jul 28, 2026

C++ implementation of a fast hash map and hash set using robin hood hashing

C++ 1,499 145 Updated Jun 13, 2026

Real-time 3D full-body reconstruction from a single camera, Multiperson BVH output, Pure C++ runtime, ONNX + ggml, 70-joint skeleton with hands.

C 619 88 Updated Jul 28, 2026

FlashKDA: high-performance Kimi Delta Attention kernels

Cuda 1,207 114 Updated Jul 30, 2026

justine's optimization for musl qsort smoothsort

C 9 Updated May 21, 2026

DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm

C 21,214 1,927 Updated Aug 9, 2026

TokenSpeed is a speed-of-light LLM inference engine.

Python 1,846 232 Updated Aug 12, 2026
Next