Skip to content
View Ankk98's full-sized avatar
💻
💻

Highlights

  • Pro

Organizations

@JIITODC

Block or report Ankk98

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

A no-fluff and highly practical masterclass that reignites engineering curiosity and helps SDE-2, SDE-3, and above become great at designing, implementing, and shipping scalable, fault-tolerant, an…

Python 3,110 640 Updated Jul 14, 2026

Distributed PostgreSQL as an extension

C 12,648 785 Updated Jul 29, 2026

NVIDIA Linux open GPU kernel module source

C 17,232 1,777 Updated Jul 7, 2026

Multi-agent deep research for Claude Code. Zero API keys. Paste one line, type /research.

JavaScript 2 Updated May 7, 2026

Professional Claude Code skills marketplace featuring production-ready skills for enhanced development workflows.

Python 1,301 210 Updated Jul 30, 2026

General-purpose deep research skill for AI agents — subagent-driven, source-backed, cited reports inline. skills.sh-compliant.

7 1 Updated May 9, 2026

A curated list of awesome skills for Cursor

Python 631 98 Updated Apr 28, 2026

Language model tokenization at GB/s

Rust 3,805 193 Updated Jul 26, 2026

Deploy autonomous AI agents as your digital twins across 10 social platforms

Handlebars 7 1 Updated Jul 17, 2026
TeX 48 8 Updated May 27, 2026

Up to 3× faster LLM decoding on Apple Silicon, lossless. Native MLX port of DeepSeek's DSpark & z-lab's DFlash speculative decoding — Gemma-4, Qwen3, Ornith-1.0, ternary Bonsai-27B.

Python 197 18 Updated Jul 26, 2026

Build a compiler to solve Anthropic's interview challenge.

Python 132 16 Updated Jul 12, 2026

Tensor library for machine learning

C++ 15,081 1,748 Updated Jul 17, 2026

LMCache: Supercharge Your LLM with the Fastest KV Cache Layer

Python 10,944 1,632 Updated Jul 30, 2026

Inferno aims to be a super lightweight, highly efficient Rust inference engine for running open weights models on Apple Silicon with Metal, targeting machines such as a MacBook Pro with 64 GB of un…

Rust 41 5 Updated Jul 29, 2026

Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

C 21,031 2,159 Updated Jul 30, 2026

LiteRT, successor to TensorFlow Lite. is Google's On-device framework for high-performance ML & GenAI deployment on edge platforms, via efficient conversion, runtime, and optimization

C++ 3,225 414 Updated Jul 30, 2026

High-efficiency floating-point neural network inference operators for mobile, server, and Web

C 2,411 534 Updated Jul 30, 2026

A close-to-metal Python API for programming AMD Ryzen™ AI NPUs (AI Engines), built on an open-source MLIR-based compiler toolchain.

C 670 197 Updated Jul 30, 2026

super repo for rocm systems projects

C++ 447 328 Updated Jul 30, 2026

Tensor library for machine learning

C++ 36 5 Updated Jul 30, 2026
Python 1,003 80 Updated Jul 29, 2026

Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.

Python 1,526 440 Updated Jul 29, 2026

A converter for transferring gguf Q4_0, Q4_1 to FLM Q4NX

Python 37 11 Updated May 27, 2026

Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.

C++ 1,655 129 Updated Jul 29, 2026

Awesome list and survey website for agents in the era of experience

178 7 Updated Jul 27, 2026

A feed-forward 3D foundation model for reconstructing scenes from streaming data

Python 15,875 1,690 Updated Jul 23, 2026

NumPy & SciPy for GPU

Python 12,221 1,121 Updated Jul 27, 2026

DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms

Python 6,810 633 Updated Jul 9, 2026

Minimal LLM inference in Rust

Rust 1,035 43 Updated Oct 24, 2024
Next