Stars
Real-time web dashboard for pi coding-agent sessions. Multi-session view, live chat mirroring, integrated terminal, diff viewer, pi-flows execution, and mobile-first remote control via mDNS or zrok…
AI inference, packed simply. A blazing-fast, zero-dependency WebGPU runtime to run GGUF models directly in the browser. Features a symmetric API for seamless local execution and cloud provider rout…
CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies
A path query language for JSON, YAML, TOML, and other serialization formats.
A modular framework for building massively parallel agentic systems
The personal finance app for everyone (by everyone)
📦️ A fast, secure MCP server that extends its capabilities through WebAssembly plugins.
NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.
AI video agents framework for next-gen video interactions and workflows.
Zero-Trust access management with true WireGuard® 2FA/MFA
Reverse Engineering: Decompiling Binary Code with Large Language Models
Build smaller, faster, and more secure desktop and mobile applications with a web frontend.
[CVPR 2024] Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data. Foundation Model for Monocular Depth Estimation
ONNX-compatible Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data
A working machine learning framework in pure Mojo 🔥
This is our own implementation of 'Layer Selective Rank Reduction'
🥧 Savoury implementation of the QUIC transport protocol and HTTP/3
Incredibly fast Whisper-large-v3
A language for constraint-guided and efficient LLM programming.
⚡ Build your chatbot within minutes on your favorite device; offer SOTA compression techniques for LLMs; run LLMs efficiently on Intel Platforms⚡
Build & ship backends without writing any infrastructure files.
[ICLR 2024 Oral] Generative Gaussian Splatting for Efficient 3D Content Creation