Skip to content
View meshkovQA's full-sized avatar

Block or report meshkovQA

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Comprehensive AI Model Evaluation Framework with advanced techniques including Temperature-Controlled Verdict Aggregation via Generalized Power Mean. Support for multiple LLM providers and 15+ eval…

Python 57 5 Updated Sep 21, 2026

A minimal, ready-to-use scaffold for launching CrewAI projects using docker-compose. Includes basic setup, configuration, and best practices to help you hit the ground running.

Python 2 4 Updated Apr 29, 2026

Полный курс по Harness 2026 на русском языке. Все что нужно знать в одщном. метсе

Shell 238 45 Updated Aug 15, 2026

Heisenbug 2026: сравнительное исследование 4 подходов к управлению контекстом AI-агентов

HTML 5 Updated Apr 23, 2026

Automated LLM testing pipeline for LM Studio using Eval AI Library. Features dynamic model loading/unloading, interactive CLI, multiple metrics (RAG, Security, Deterministic), and integrated web da…

Python 2 Updated Mar 17, 2026
Python 7 7 Updated Mar 25, 2026

Evaluation framework for agentic AI systems with CI/CD-enforced quality gates, hallucination detection, and tool-call validation

Python 2 1 Updated Mar 9, 2026

WebSocket testing toolkit — CLI + Web Dashboard. Connect, record & replay, load test, validate schemas.

TypeScript 12 Updated Feb 19, 2026

Metrics for Kaggle competitions 📏

Python 9 3 Updated Jul 19, 2019