Skip to content
View krk's full-sized avatar

Block or report krk

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

Python 18,972 1,483 Updated Jul 23, 2026

Pure Rust Inference Engine

Rust 609 86 Updated Jul 24, 2026

Source code for the Microsoft Comic Chat IRC client

C++ 958 108 Updated Jul 22, 2026

Import OpenStreetMap data into a PostgreSQL/PostGIS database

C++ 1,674 481 Updated Jul 5, 2026

Library for shoehorning the Slug text/graphics GPU rendering library into projects.

C++ 131 7 Updated Jul 19, 2026

GLM-5.2 Quantrio INT4/INT8 Mixed Abliterated — SPEED=1 C1≈30 tok/s @ 128k on 4x DGX Spark. Step-by-step recipe, image bake, results.

Python 12 1 Updated Jul 18, 2026

Postgres rewritten in Rust, now passing 100% of the Postgres regression tests

Rust 3,750 138 Updated Jul 10, 2026

Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

C 18,416 1,787 Updated Jul 24, 2026

Open Source search based on OpenStreetMap data

Python 4,390 845 Updated Jul 8, 2026

MiniMax-M3 (428B, no pruning) at 36 tok/s on 2× NVIDIA DGX Spark — W4A16 GPTQ + NVFP4 KV + EAGLE-3 speculative decoding on vLLM. Three serving lanes: speed / balanced / long-context.

39 3 Updated Jul 13, 2026

DeepSeek V4 Flash DSpark 1M NVFP4 KV recipe for 2x DGX Spark

Python 155 11 Updated Jul 16, 2026

Docker configuration for running VLLM on dual DGX Sparks

Shell 1,884 344 Updated Jul 24, 2026

GLM-5.2 (744B/40B MoE) on a 4× DGX Spark / GB10 (sm_121) cluster: portable Triton sparse-MLA kernels, a data-free expert prune, MTP draft, and a one-script bootstrap.

Python 7 Updated Jun 24, 2026

MiMo-V2.5 Omni TP=2 on 2x DGX Spark · 1M context · NVFP4 4-bit KV (~1.97M-token KV pool @ 1M, ~30 tok/s) · 69-eval: thinking-OFF 97.8 beats thinking-ON 90.6 for tool/agent work

Python 37 4 Updated Jul 13, 2026

turn qwen 3.6 27B-> AR -> diffusion (opendllm + d3llm)

Python 13 Updated Jun 11, 2026

A Docker-based pipeline for fine-tuning FLUX.1-dev with LoRA on the DGX Spark

Python 6 Updated Dec 3, 2025

NVIDIA DGX Spark Playbooks

Python 18 3 Updated Nov 26, 2025

This is the official code of the paper "A Multi-Agent System Enables Versatile Information Extraction from the Chemical Literature"

Python 106 20 Updated Jun 16, 2026

This is the official code of the paper "MolNexTR: a generalized deep learning model for molecular image recognition"

Jupyter Notebook 168 37 Updated Jan 15, 2026

MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.

Python 5,248 678 Updated Jul 24, 2026

Convert PDF to markdown + JSON quickly with high accuracy

Python 37,804 2,655 Updated Jul 20, 2026

Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.

Python 18,355 1,742 Updated Jul 23, 2026

DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm

C 19,136 1,677 Updated Jul 23, 2026

advanced compilers

HTML 981 226 Updated Jan 10, 2026

🪅 Windows & Linux userspace emulator

C++ 3,424 216 Updated Jul 22, 2026

A JavaScript library aimed at visualizing graphs of thousands of nodes and edges

TypeScript 12,111 1,624 Updated Jul 8, 2026

A 5-20x faster experimental Homebrew alternative

Rust 7,486 174 Updated Jun 12, 2026

The fastest macOS package manager. Written in Zig. 3ms warm installs.

Zig 1,100 17 Updated Jul 21, 2026
Next