Highlights
-
-
terminal-bench-science Public archive
Forked from harbor-framework/terminal-bench-scienceTerminal-Bench Science: Evaluating AI Agents on Complex Real-World Scientific Workflows in the Terminal
-
schneidergithub.github.io Public archive
Aaron Schneider's personal website.
-
-
harbor-benchmark-template Public template
Forked from harbor-framework/benchmark-templateBenchmarks are Software
-
fedramp-automation Public archive
Forked from FedRAMP/rulesThis repository contains the structured machine-readable rules for FedRAMP.
-
model-runner Public
Forked from docker/model-runnerDocker Model Runner
-
mini-swe-agent Public archive
Forked from SWE-agent/mini-swe-agentThe 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench verified!
-
chess-experiment-ghcopilot-opus4.8 Public archive
my own benchmarking test
-
node-ipc Public archive
Forked from RIAEvangelist/node-ipcA nodejs module for local and remote Inter Process Communication (IPC), Neural Networking, and able to facilitate machine learning.
-
-
-
-
ml-learning-path Public
Forked from Jay06eng/ml-learning-pathA structured, hands-on ML learning path from statistics fundamentals to applied algorithms — Python, Scikit-learn, and Google Colab.
-
pmac Public archive
Project Management as Code
-
openai-evals Public archive
Forked from openai/evalsEvals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
-
webmcp Public archive
Forked from webmachinelearning/webmcp🤖 WebMCP
-
docs Public archive
Forked from github/docsThe open-source repo for docs.github.com
-
eureka-ml-insights Public archive
Forked from microsoft/eureka-ml-insightsA framework for standardizing evaluations of large foundation models, beyond single-score reporting and rankings.
-
MMMU Public archive
Forked from MMMU-Benchmark/MMMUThis repo contains evaluation code for the paper "MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI"
-
outline Public archive
Forked from outline/outlineThe fastest knowledge base for growing teams. Beautiful, realtime collaborative, feature packed, and markdown compatible.
-
toolkit Public archive
Forked from actions/toolkitThe GitHub ToolKit for developing GitHub Actions.
-
copilot-sdk Public archive
Forked from github/copilot-sdkMulti-platform SDK for integrating GitHub Copilot Agent into apps and services
TypeScript MIT License UpdatedJan 22, 2026 -
github-pages Public archive
Forked from skills/github-pagesCreate a site or blog from your GitHub repositories with GitHub Pages.
-
Git Source Code Mirror - This is a publish-only repository but pull requests can be turned into patches to the mailing list via GitGitGadget (https://gitgitgadget.github.io/). Please follow Documen…
-
BenchmarkTransferLearning Public archive
Forked from MR-HosseinzadehTaher/BenchmarkTransferLearningOfficial PyTorch Implementation and Pre-trained Models for Benchmarking Transfer Learning for Medical Image Analysis
-
dots.vlm1 Public archive
Forked from studio-dots-ai/dots.vlm1The official repository of the dots.vlm1 instruct models proposed by rednote-hilab.
-
K2-Think-SFT Public archive
Forked from MBZUAI-IFM/K2-Think-SFT -
agentic-misalignment Public
Forked from anthropic-experimental/agentic-misalignmentPython MIT License UpdatedJun 19, 2025 -