Git.Top Guide

Git.Top Score Guide

Understand Git.Top Score, Agent Score, quality confidence, and the signals behind project selection.

Guide

What The Score Means

Git.Top Score summarizes maintenance, community, documentation, stability, adoption, deployment fit, and agent readability into a decision aid instead of a popularity counter.

Guide

How To Use It

Inspect the score breakdown, confidence, freshness, and quality evidence before comparing projects or sending an agent-readable recommendation.

Data Source

d1 / d1_query

1213 loaded projects. Generated at 2026-09-26T15:44:08.468Z.

Evaluation package that allows benchmarking of agentic AIs from various sources and frameworks by producing statistical results which can be compared across different use cases and datasets.

Llm EvalLibrary OnlyLocalCloud
Quality6
Agent53

A Model Context Protocol (MCP) server that provides programmatic control over MuseScore!

Mcp ServerLibrary OnlyLocalCloud
Quality6
Agent54

Project

modu-ai/moai-adk

67

Agentic development harness for Claude Code — SPEC-driven plan/run/sync, TRUST 5 quality gates, model+effort routing, and Claude×GLM multi-LLM cost control. Single Go binary, 16 languages, zero deps.

Coding AgentLocalCloud
Quality32
Agent67

Project

qdrant/skills

62

Agent skills for Qdrant vector search: scaling, performance optimization, search quality, monitoring, deployment, model migration, version upgrades, and SDK usage across Python, TypeScript, Rust, Go, .NET, Java

Ai ObservabilityKubernetesLocalCloud
Quality15
Agent62

📚 A curated list of papers & technical articles on AI Quality & Safety

Llm EvalLibrary OnlyLocalCloud
Quality5
Agent54

Development workflows for Claude Code that keep broad exploration focused on the outcome you approved.

Workflow AutomationLibrary OnlyLocalCloud
Quality25
Agent65

Evidently is ​​an open-source ML and LLM observability framework. Evaluate, test, and monitor any AI-powered system or data pipeline. From tabular data to Gen AI. 100+ metrics.

Ai ObservabilityDockerLibrary OnlyLocal
Quality24
Agent76

Official Monte Carlo toolkit for AI coding agents. Skills and plugins that bring data and agent observability — monitoring, triaging, troubleshooting, health checks — into Claude Code, Cursor, and more.

Mcp ServerLocalCloud
Quality23
Agent62

A Python client to interact with Arize API

Ai ObservabilityLibrary OnlyLocalCloud
Quality22
Agent62

A framework for standardizing evaluations of large foundation models, beyond single-score reporting and rankings.

Llm EvalLibrary OnlyLocalCloud
Quality8
Agent57

EvalBench is a flexible framework designed to measure the quality of generative AI (GenAI) workflows around database specific tasks.

Llm EvalDockerLocalCloud
Quality36
Agent71