Skip to content
View wwymak's full-sized avatar

Highlights

  • Pro

Organizations

@CoronaWhy

Block or report wwymak

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Skills for building LLM evals. Starts with interactive error discovery: build a review app, sample diverse traces, and organize human annotations into failure modes.

224 19 Updated Aug 16, 2026

Switchyard lets LLM applications route traffic across models and providers while preserving native OpenAI and Anthropic API compatibility - enabling flexible model selection, benchmarking, and cost…

Rust 1,757 154 Updated Aug 17, 2026

Simulate Before Reality.

Python 1,152 171 Updated Aug 16, 2026

ADR secures enterprise AI agents through observability, security benchmarking, and threat detection. Deployed at Uber.

Python 1,442 127 Updated Aug 16, 2026

A self-improving RLM agent for coding workflows and long-running autonomous tasks.

TypeScript 16,752 1,801 Updated Aug 17, 2026

Programmatic memory for long-horizon LLM agents: the harness appends everything to one log, and the agent searches it with code. 97.4% on ARC-AGI-3 (arXiv:2607.20064)

Python 399 40 Updated Aug 15, 2026

Multiplayer agent harness for work.

TypeScript 13,751 1,623 Updated Aug 17, 2026

Standards for defining and evaluating agent behavior

TypeScript 275 6 Updated Jul 29, 2026

Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.

Python 22,409 2,361 Updated Aug 13, 2026

Local speech-to-text for macOS on-device AI, fully private, optional cloud

Swift 1,696 121 Updated Aug 17, 2026
Python 238 26 Updated Jun 8, 2026

Fully automatic censorship removal for language models

Python 27,748 3,002 Updated Aug 17, 2026

⌥ AI Coding agent for the terminal — hash-anchored edits, optimized tool harness, LSP, Python, browser, subagents, and more

TypeScript 25,362 2,444 Updated Aug 17, 2026

AutoMem: Automated Learning of Memory as a Cognitive Skill

Python 135 17 Updated Jul 3, 2026

A Python port of Pi’s minimalist coding agent.

Python 2,346 279 Updated Aug 17, 2026

Companion repo to our livestream series Show Us Your (Agent) Skills

JavaScript 64 9 Updated Aug 12, 2026

Long-term memory for AI assistants. Graph + vector store that recalls decisions, relationships, and context across sessions.

Python 799 101 Updated Aug 14, 2026

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,922 1,002 Updated Aug 13, 2026

Bridging LLM and Recommender System.

Jupyter Notebook 1,187 127 Updated Jan 27, 2026

Security scanner for AI agent skills. Detect vulnerabilities, malicious patterns, security risks, prompt injection, data exfiltration, and supply-chain risks in Claude Code, Codex, and MCP skills b…

Python 14,695 1,224 Updated Aug 16, 2026

A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes…

Shell 145,913 23,580 Updated Aug 6, 2026

An interface library for RL post training with environments.

Python 2,503 427 Updated Aug 13, 2026

Synthetic Data Generation Toolkit for LLMs

Python 157 61 Updated Aug 12, 2026

the LLM vulnerability scanner

Python 8,834 1,175 Updated Aug 14, 2026

The absolute trainer to light up AI agents.

Python 17,490 1,541 Updated Aug 17, 2026

Ingest, parse, and optimize any data format ➡️ from documents to multimedia ➡️ for enhanced compatibility with GenAI frameworks

Python 7,811 665 Updated Dec 12, 2025

#1 Persistent memory for AI coding agents based on real-world benchmarks

TypeScript 27,101 2,313 Updated Aug 17, 2026
Next