Starred repositories
Distributed AI/LLM for the people. Share compute privately or publicly to power your agents and chat.
A 2.78-trillion-parameter Kimi K3 running inference on a single CPU in 8.24 GB of RAM. Portable C99: no BLAS, no framework, no GPU.
ICML 2026 · Plug-and-play long-term memory for LLM agents
📝A simple and elegant markdown editor, available for Linux, macOS and Windows.
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
Truthful scheduler-step and streaming visualization for local LLM inference
AirLLM 70B inference with single 4GB GPU
Run the full 2.78-trillion-parameter Kimi K3 model beyond available RAM by streaming activated weights directly from NVMe. A dependency-free, embeddable C inference engine.
Local UI to run and train LLMs and diffusion models, including Kimi K3, MiniMax-H3, Gemma 4, Qwen3.6, DeepSeek-V4, FLUX and more.
A curated list of public-source, research, and commercial tools for AI security and AI-assisted cybersecurity — autotriage, agent security, AI/ML supply chain, pentest agents, AI SAST, LLM-driven f…
[NeurIPS 2025] MemEIC: A Step Toward Continual and Compositional Knowledge Editing
Bootable Model As System (BMASS) – an offline-first AI operating environment
π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.
Code at the speed of thought – Zed is a high-performance, multiplayer code editor from the creators of Atom and Tree-sitter.
The open-source AI voice studio. Clone, dictate, create.
A collection of pragmatic, real-world examples guiding you from basic to advanced use of xAI's Grok APIs.
Fast local transcription for large lectures with NVIDIA Parakeet ONNX
Private migration backup for cashfromchaos
Collection of skills for small businesses using AI Agents
Kernel-enforced authority and spend platform for AI agents
SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.
🤱🏻 Turn any webpage into a desktop app with one command.
Light, fluffy, and always free - The AWS Local Emulator alternative
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
Headless client for Obsidian Sync. Sync your vaults from the command line without the desktop app.
An agentic skills framework & software development methodology that works.