Skip to content
#

ai-metrics

Here are 23 public repositories matching this topic...

AgentMeasure

Open measurement infrastructure for AI agents. Our audit of 124 usage tools found 45+ verified billing bugs — 19 fixes landed upstream. Conformance fixtures for token accounting: PASS / FAIL / UNPROVABLE in CI.

  • Updated Sep 22, 2026
  • Python

Agentic Workflow Evaluation: Text Summarization Agent. This project includes an AI agent evaluation workflow using a text summarization model with OpenAI API and Transformers library. It follows an iterative approach: generate summaries, analyze metrics, adjust parameters, and retest to refine AI agents for accuracy, readability, and performance.

  • Updated Feb 23, 2025
  • Python

Add this topic to your repo

To associate your repository with the ai-metrics topic, visit your repo's landing page and select "manage topics."

Learn more