Git.Top Guide

Git.Top Score Guide

Understand Git.Top Score, Agent Score, quality confidence, and the signals behind project selection.

Guide

What The Score Means

Git.Top Score summarizes maintenance, community, documentation, stability, adoption, deployment fit, and agent readability into a decision aid instead of a popularity counter.

Guide

How To Use It

Inspect the score breakdown, confidence, freshness, and quality evidence before comparing projects or sending an agent-readable recommendation.

Data Source

d1 / d1_query

1067 loaded projects. Generated at 2026-08-13T13:33:56.552Z.

Evaluation package that allows benchmarking of agentic AIs from various sources and frameworks by producing statistical results which can be compared across different use cases and datasets.

Llm EvalLibrary OnlyLocalCloud
Quality6
Agent53

A Model Context Protocol (MCP) server that provides programmatic control over MuseScore!

Mcp ServerLibrary OnlyLocalCloud
Quality5
Agent52

Project

modu-ai/moai-adk

67

Agentic development harness for Claude Code — SPEC-driven plan/run/sync, TRUST 5 quality gates, model+effort routing, and Claude×GLM multi-LLM cost control. Single Go binary, 16 languages, zero deps.

Coding AgentLocalCloud
Quality31
Agent67

Project

qdrant/skills

60

Agent skills for Qdrant vector search: scaling, performance optimization, search quality, monitoring, deployment, model migration, version upgrades, and SDK usage across Python, TypeScript, Rust, Go, .NET, Java

Ai ObservabilityLocalCloud
Quality15
Agent60

📚 A curated list of papers & technical articles on AI Quality & Safety

Llm EvalLibrary OnlyLocalCloud
Quality5
Agent54

Development workflows for Claude Code that keep broad exploration focused on the outcome you approved.

Workflow AutomationLibrary OnlyLocalCloud
Quality30
Agent67

Evidently is ​​an open-source ML and LLM observability framework. Evaluate, test, and monitor any AI-powered system or data pipeline. From tabular data to Gen AI. 100+ metrics.

Ai ObservabilityDockerLibrary OnlyLocal
Quality24
Agent76

A Python client to interact with Arize API

Ai ObservabilityLibrary OnlyLocalCloud
Quality10
Agent57

A framework for standardizing evaluations of large foundation models, beyond single-score reporting and rankings.

Llm EvalLibrary OnlyLocalCloud
Quality8
Agent57

EvalBench is a flexible framework designed to measure the quality of generative AI (GenAI) workflows around database specific tasks.

Llm EvalDockerLocalCloud
Quality35
Agent70

LLM evaluation framework for Elixir: evaluate and test LLM outputs, detect hallucinations, measure response quality

Llm EvalLocalCloud
Quality17
Agent56

OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards

Llm EvalLibrary OnlyLocalCloud
Quality9
Agent60