Lists (3)
Sort Name ascending (A-Z)
Stars
A self-improving RLM agent for coding workflows and long-running autonomous tasks.
EnterpriseClawBench: Benchmarking Agents from Real Workplace Sessions
SpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
cut Claude Code token usage by rendering text context as images
Scaling the Horizon, Not the Parameters
Research Pilot helps AI agents understand, maintain, and advance research projects through papers, claims, evidence, experiments, and human-gated updates.
Enhanced LanceDB memory plugin for OpenClaw — Hybrid Retrieval (Vector + BM25), Cross-Encoder Rerank, Multi-Scope Isolation, Management CLI
Profile-driven autonomous kernel optimization loop — a portable skill file for AI agents (CUDA/Triton/ROCm/CPU/SIMD)
Autonomous no-human-in-the-loop optimization research for Claude Code
AI agents running research on single-GPU nanochat training automatically
Programmatic video for coding agents — HTML to video on your laptop. Turn HTML, CSS & data into real MP4s with pluggable render engines, 21 templates, AI soundtrack. Apache-2.0, no per-render fees.…
An open, collaboratively-built repository for AI-assisted scientific research — collecting and curating agents, skills, workflows, tools, and best practices across the full research lifecycle. 面向 A…
Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
Reference code for the Meta-Harness paper.
AI turns documents or topics into real, native PowerPoint decks—with native shapes, transitions and animations, data-backed charts and tables on demand, audio narration from speaker notes, and supp…
Agent skill for running OpenSpec SDD end-to-end without human-in-the-loop next prompts.