Popular repositories Loading
-
qed-bench
qed-bench PublicReproducible benchmarks comparing small scoring models against LLM-as-judge on essay quality, spam, AI-text detection, and LLM authorship โ measuring quality, cost, and latency on the same Pareto fโฆ
Jupyter Notebook
-
Repositories
- docs Public
- qed-bench Public
Reproducible benchmarks comparing small scoring models against LLM-as-judge on essay quality, spam, AI-text detection, and LLM authorship โ measuring quality, cost, and latency on the same Pareto frontier.
- plugins Public
Plugins (Claude Code, GitHub Actions) for scoring, evaluating, and improving content with DLMs (U/=22A8)
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Top languages
Loadingโฆ
Most used topics
Loadingโฆ