Skip to content
View PieBru's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report PieBru

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

RTX 4090 (sm_89) Linux port of NInfer for Qwen3.6-27B

C++ 8 Updated Aug 2, 2026

High-performance single-GPU inference for selected model checkpoints and GPUs.

C++ 759 114 Updated Aug 19, 2026

Meta-Framework of Spatiotemporal Composability

TypeScript 6,264 346 Updated Aug 13, 2026

🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

Python 34,342 7,249 Updated Aug 19, 2026

DeepSeek Harness: Everything is a Plugin.

TypeScript 163,942 17,351 Updated Aug 17, 2026

Pinned one-DGX-Spark Docker recipe for DeepSeek V4 Flash with EXL3, SparkInfer, and 262K NVFP4 MLA KV cache

Python 27 4 Updated Aug 11, 2026

ROCmFPX Family for AMD Hardware and Processors. More quants and special agent quants

C++ 267 40 Updated Aug 17, 2026

Performance-tuned llama.cpp for AMD Strix Halo (gfx1151): FA + MoE-prefill fixes with a bundled current Mesa driver. Vulkan and HIP; portable dir, Docker, and distrobox.

Python 79 11 Updated Aug 19, 2026

Run huge MoE models from a swarm of peers, with the colibri engine. Pure C.

C 117 8 Updated Aug 18, 2026

Local-first voice-to-text for pi, using Whisper-style backends to dictate coding prompts directly into the editor.

TypeScript 7 1 Updated Aug 10, 2026

Monorepo for @juicesharp/rpiv-* Pi plugins (lockstep versions, single install, single publish pipeline)

TypeScript 636 113 Updated Aug 18, 2026

Safe push-to-talk dictation for Pi with pluggable transcription backends

JavaScript 1 Updated Aug 18, 2026

OCR model that handles complex tables, forms, handwriting with full layout.

Python 12,119 1,232 Updated Jun 26, 2026

a community oriented 1:1, vLLM-alike (Continuous batching, paged KV) engine in C++ with additional features (for example, RadixAttention, Cache-aware scheduling)

C++ 306 43 Updated Aug 19, 2026

TypeScript AI agent framework where personality is architecture. Specialists, not chatbots.

TypeScript 16 Updated Aug 18, 2026

A hive mind communication platform

Rust 28,488 3,555 Updated Aug 19, 2026

Persistent project memory for AI coding agents. Structured scaffold + drift detection CLI.

TypeScript 1,489 103 Updated Aug 18, 2026

A full featured, powerful, and efficient AI Agentic Harness/Desktop Agent app designed from the ground up with local inference on consumer hardware in mind. It evolves, grows, and gets smarter as y…

Python 111 11 Updated Aug 19, 2026

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.

203,742 20,892 Updated Apr 20, 2026

Repeatable agents-plus-code workflows, packaged as one skill, stamped into any repo. Deterministic Python owns the graph; coding agents are bounded nodes inside it.

Python 697 165 Updated Aug 4, 2026

Mixxx is Free DJ software that gives you everything you need to perform live mixes.

C++ 7,053 1,847 Updated Aug 18, 2026

Agent skills synthesizing years of software engineering discipline into a prescriptive methodology for solo developers

Shell 144 9 Updated Aug 7, 2026

TencentDB Agent Memory is a team-level memory hub for AI Agents — turning conversations, docs, and code into four reusable memory assets (Chat Memory, Skill, LLM-Wiki, Code-Graph) that are governed…

TypeScript 23,107 2,112 Updated Aug 15, 2026

An Open Source Python alternative to NotebookLM's podcast feature: Transforming Multimodal Content into Captivating Multilingual Audio Conversations with GenAI

Python 6,500 760 Updated May 4, 2026

A 2.78-trillion-parameter Kimi K3 running inference on a single CPU in 8.24 GB of RAM. Portable C99: no BLAS, no framework, no GPU.

C 6,050 992 Updated Aug 7, 2026

Run the full 2.78-trillion-parameter Kimi K3 model beyond available RAM by streaming activated weights directly from NVMe. A dependency-free, embeddable C inference engine.

C 2,198 163 Updated Aug 13, 2026

Open-source observability for your GenAI or LLM application, based on OpenTelemetry

Python 7,384 1,053 Updated Aug 10, 2026

SenseNova-U series: Native Unified Paradigm with NEO-unify from the First Principles

Python 4,953 416 Updated Aug 19, 2026

The most RAM efficient harness

Rust 17,969 2,019 Updated Aug 19, 2026
Next