router_benchmark
PythonA common interface evaluation harness for comparing agentic LLM routers (LiteLLM, RouteLLM, vLLM Semantic Router) across live benchmarks. Scores on a deployment focused metric suite (cost, latency, tool call accuracy, Pareto frontier) instead of a single leaderboard number.
View on GitHub →