Stars
A persistent workspace for development work that self-improves and continues beyond one session.
edwinbrowwn / llama.cpp-rdna2
Forked from ggml-org/llama.cppLLM inference in C/C++
Privacy-first AI meeting notes & writing assistant for Mac — on-device transcription, smart minutes, action items & 60+ writing modes, running entirely on Apple Silicon
A curated list of free, legitimate AI/ML books
A tool to unlobotomize your NVIDIA card!
JavaScript in-page GUI agent. Control web interfaces with natural language.
carlosfundora / sglang-1-bit-turbo
Forked from sgl-project/sglangAMD ROCm (gfx1030) inference fork with RotorQuant/TurboQuant KV compression, PHANTOM-X zero-copy draft speculation, EAGLE3 speculative decoding, 12 RDNA2 crash fixes, and PrismML Bonsai Q1_0_G128 1…
A lightweight, lightning-fast, in-process vector database
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
The world's most powerful open-source bio AI assistant - Access academic literature, clinical trials, drug labels, and more, all through natural conversation.
30 MCP tools for local AI agent swarms - runs entirely on qwen3:4b
iacopPBK / llama.cpp-gfx906
Forked from ggml-org/llama.cppllama.cpp-gfx906
A Python library for extracting structured information from unstructured text using LLMs with precise source grounding and interactive visualization.
Extract structured data from PDFs, Word docs and images. Embeddable directly into your application, regardless of the stack.
pdfLLM is a completely open source, proof of concept RAG app.