Highlights
Lists (1)
Sort Name ascending (A-Z)
Starred repositories
Multi-slot LLM inference on AMD Strix Halo: recipes + honest benchmarks (236 tok/s @ 32 streams, llama.cpp Vulkan)
A list of Free Software network services and web applications which can be hosted on your own servers
Locally fine-tuned Russian RAG models: document splitter, query expansion, retrieval embedder. Teacher distillation → GGUF → llama.cpp on AMD Vulkan. Commercial-OK licenses.
Private LLM/RAG platform in one command for NVIDIA DGX Spark / GB10 (arm64). Validated on real hardware.
One-command local AI/RAG installer for macOS (Metal): Dify, Open WebUI, Ollama, Weaviate/Qdrant, Postgres, Redis. 230 tests.
WAKE.md for AI agents: compile project state so agents stop starting cold.
Self-hosted LLM/RAG stack in one command — AMD Strix Halo / x86_64 (ROCm/Vulkan, Docker Compose)