Skip to content
View piDack's full-sized avatar
🎯
Focusing
🎯
Focusing
  • Beijing
  • 15:33 (UTC +08:00)

Block or report piDack

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

FlashKDA: high-performance Kimi Delta Attention kernels

Cuda 466 47 Updated May 26, 2026

CUDA kernels for linear attention variants, written in CuTe DSL and CUTLASS C++.

Python 535 70 Updated Jul 23, 2026

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.

Rust 194,880 109,514 Updated Jun 26, 2026

Fast and memory-efficient exact attention

Python 24,522 2,933 Updated Jul 23, 2026

🚀 Efficient implementations for emerging model architectures

Python 5,405 605 Updated Jul 24, 2026

Open-source, desktop-grade AI agent that gets real work done — data analysis, slides, docs, video & web research. Built on OpenClaw; runs tools on your real desktop and takes commands from your pho…

TypeScript 5,663 885 Updated Jul 24, 2026

Tiny, Fast, and Deployable anywhere — automate the mundane, unleash your creativity

Go 29,709 4,316 Updated Jul 23, 2026

Lightweight, open-source AI agent for your tools, chats, and workflows.

Python 46,161 8,165 Updated Jul 24, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 383,974 80,670 Updated Jul 24, 2026

Youtu-Tip: Tap for Intelligence, Keep on Device.

Python 591 66 Updated Feb 27, 2026

On-device AI across mobile, embedded and edge for PyTorch

Python 4,825 1,085 Updated Jul 24, 2026

Tile-Based Runtime for Ultra-Low-Latency LLM Inference

Python 1,586 111 Updated Jul 14, 2026

Native cross-platform system automation

C++ 220 41 Updated Dec 13, 2022

Mouse, keyboard and display automation.

C++ 26 2 Updated Jul 5, 2026

Build Conversational AI in minutes ⚡️

Python 12,325 1,728 Updated Jun 11, 2026

An Open Phone Agent Model & Framework. Unlocking the AI Phone for Everyone

Python 25,847 4,011 Updated Mar 6, 2026

MoBA: Mixture of Block Attention for Long-Context LLMs

Python 2,151 157 Updated Apr 3, 2025

AHN: Artificial Hippocampus Networks for Efficient Long-Context Modeling

Python 181 5 Updated Oct 17, 2025

The Company AI Command Center

TypeScript 20,027 3,431 Updated Jul 24, 2026

⚙️ Create and run workflows (RPA 2.0)

Python 4,104 335 Updated Jul 17, 2026

Examples of CUDA implementations by Cutlass CuTe

Makefile 280 35 Updated Jul 1, 2025
C++ 9 Updated Sep 11, 2025

An open-source framework for detecting, redacting, masking, and anonymizing sensitive data (PII) across text, images, and structured data. Supports NLP, pattern matching, and customizable pipelines.

Python 10,150 1,219 Updated Jul 24, 2026

MCP Server for Computer Use in Windows

Python 6,490 790 Updated Jul 20, 2026

🐉 Revolutionary NPU framework for Linux | 24,988 FPS face recognition | AMD XDNA support | World's first complete NPU stack

Python 44 5 Updated Aug 7, 2025

Run frontier LLMs and VLMs locally on Qualcomm devices across NPU, GPU, and CPU with a few lines of code

Rust 8,251 1,030 Updated Jul 24, 2026

An open-source coding helper. Very friendly!

TypeScript 993 76 Updated Jul 24, 2026

Train speculative decoding models effortlessly and port them smoothly to SGLang serving.

Python 1,005 292 Updated Jul 24, 2026

GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models

Python 4,404 468 Updated Feb 1, 2026

Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.

C++ 1,646 125 Updated Jul 23, 2026
Next