Stars
Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including CrewAI, Agno, OpenAI Agents SDK, Langchain, Autogen, AG2, andโฆ
A self-improving skill for AI coding agents (Claude Code, Cursor, AGENTS.md): recognize a hard-won golden path in a session and harvest it into a reusable skill/rule for next time.
A tool for creating and running Linux containers using lightweight virtual machines on a Mac. It is written in Swift, and optimized for Apple silicon.
Framework for evaluating and improving agents
Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
Start a run in your terminal and walk away. Get pinged when it finishes, or needs you. Every step is a readable trace you check before anything ships.
Public ant-irys code and Harvey LAB benchmark results
Continuous background code review database for agents, work faster and smarter with accountability for every line of generated code.
Open-source & free โ Battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in fine-tuned ruleset (NPE, threโฆ
Skills for threat modeling, scanning, triage, patching, plus an autonomous scanning harness you can /customize
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
Free and open source database client built natively for developers
Multi-harness agentic plugin marketplace for Claude Code, Codex CLI, Cursor, OpenCode, GitHub Copilot, and Gemini CLI
A GPU-accelerated cross-platform terminal emulator and multiplexer written by @wez and implemented in Rust
AutoEvals is a tool for quickly and easily evaluating AI model outputs using best practices.
AI-powered QA testing framework that uses LLMs (Claude or GPT) to test web apps, CLI tools, and TUI programs from markdown story cards, returning structured pass/fail verdicts with evidence.
A lightweight alternative to OpenClaw that runs in containers for security. Connects to WhatsApp, Telegram, Slack, Discord, Gmail and other messaging apps,, has memory, scheduled jobs, and runs dirโฆ
Open-source credential gateway with a built-in vault. give your AI agents access to services without exposing keys.
LLM-supervised persistent memory for AI agents โ graph-based recall, cross-session knowledge, single binary. Works with Claude Code, OpenClaw, and any CLI agent.
agent multiplexer that lives in your terminal.
Lightweight and Memory efficient terminal for Mac built with SwiftUI and libghostty
A native macOS terminal for agent-driven development, built on Ghostty.
SwiftUI component for displaying rich release notes inside an app
The headless browser for AI agents and web scraping
๐น Clean, uninstall, analyze, optimize, and monitor your Mac from the terminal.
๐ A fast, out-of-the-box terminal built for AI coding.
๐๐ผโโ๏ธ The blackboard for coding agents - multi-session tool for claude code, cursor, codex, gemini