Stars
Agent workspace built on Cloudflare Workers for creating documents, building apps, and running agents with your company’s context and systems.
A self-improving RLM agent for coding workflows and long-running autonomous tasks.
AirLLM 70B inference with single 4GB GPU
Universal SEO skill for Claude Code. 25 sub-skills + 18 sub-agents covering technical SEO, E-E-A-T, schema, GEO/AEO, backlinks, local SEO, maps intelligence, semantic clustering, e-commerce SEO, in…
The AI agent sandbox runtime. Boot any OCI image as a Firecracker microVM, Linux container, or Apple Silicon VM in ~400 ms — zero-copy DAG storage, P2P image sync, native MCP for opencode/Claude Co…
Point it at any web page and it finds the video, extracts the stream, transcodes it and casts in real time to your TV. It even burns subtitles….
Use your NVIDIA GPU's VRAM as swap space on Linux. Built for laptops with soldered memory and no upgrade path. If you have an RTX card sitting there with 8GB of VRAM and you're getting swapped to S…
Open source, cloud native, Postgres platform with copy-on-write branching and scale-to-zero
Snap any video URL or audio file into plaintext. No GPU. No cloud. One command.
AI watermark remover. CLI and Python library to strip visible and invisible AI watermarks (Gemini / Nano Banana sparkle, SynthID) and provenance metadata (C2PA, EXIF, IPTC) from images.
An agentic skills framework & software development methodology that works.
Ping a Neon Serverless Postgres database using a Vercel Edge Function to see the journey your request makes.
powerful claude code wrapper - arbitrary message branching & tab system, remote-control, chat search, drafts and more
⚡ LFK is a lightning-fast, keyboard-focused, yazi-inspired terminal user interface for navigating and managing Kubernetes clusters. Built for speed and efficiency, it brings a three-column Miller c…
Achieve state of the art inference performance with modern accelerators on Kubernetes
Lightweight HLS restream proxy for Jellyfin/Emby/Plex — injects headers, rewrites playlists, auto-refreshes tokens
Library for reducing tail latency in RAM reads
TurboQuant: Near-optimal KV cache quantization for LLM inference (3-bit keys, 2-bit values) with Triton kernels + vLLM integration
A plain-text memory system for AI agents. Clone it, point your agent at it, start remembering.
Running a big model on a small laptop
SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log m…
🔐 Authentication, Authorization, and Accounting (AAA) App and Plugin for Caddy v2. 💎 Implements Form-Based, Basic, Local, LDAP, OpenID Connect, OAuth 2.0 (Github, Google, Facebook, Okta, etc.), SAM…
Open-source CUDA, Triton and HIP compiler targeting multiple GPU and CPU architectures.
Visualize and share your data. All in SQL. Powered by DuckDB.
The most popular Kubernetes Operator for PostgreSQL.