Project Knowledge

supermemoryai/memorybench

Use supermemoryai/memorybench when the user needs a llm eval project with local, cloud deployment options.

llm_evalaiai-memorybenchmark-framework
TypeProject / implementation
Difficultybeginner
LanguageTypeScript
LicenseMIT

TL;DR

Use supermemoryai/memorybench when the user needs a llm eval project with local, cloud deployment options.

Use supermemoryai/memorybench when the user needs a llm eval project with local, cloud deployment options.

Install

Install and run the evaluation harness as documented by the repository.

LocalCloud

Good For

evaluate LLM outputs
benchmark prompts and agents
track model quality

Not Good For

edge-only Cloudflare Workers deployment without adaptation
production inference serving
end-user chat apps
Git.Top Score69/100
Agent Score55/100
Maintenance11
Stability70
Quality6/100

Best Use

evaluate LLM outputs

Primary situation where this project is a good shortlist candidate.

Watch Out

edge-only Cloudflare Workers deployment without adaptation

Main reason to compare alternatives before adopting it.

Deployment Fit

local, cloud

Git.Top did not classify this as Cloudflare-ready. Evidence: No Cloudflare deployment signal detected.

Best Alternative

promptfoo/promptfoo: docker, library_only

Compare this option when the target stack or language preference differs.

Confidence

2/4 classification signals high; 4 quality signals complete or snapshot

Classification evidence and quality signal confidence should be checked before production recommendations.

Freshness

Repo 2026-08-04; metrics 2026-08-04

Repository and metric timestamps for this knowledge record.

Freshness

Repository and metrics timestamps

Repository synced2026-08-04T13:30:20.878Z
Metrics calculated2026-08-04T13:30:20.878Z

Scoring

Quality and agent score are separate

Quality score weights star movement, commits, releases, contributors, and issue response. Agent score weights documentation, maintenance, deployment, popularity, and community.

Confidence

Signal confidence

Stars 30dsnapshot
Commits 30dcomplete
Releases 180dcomplete
Contributors 90dcomplete

Badge

Agent Score badge

Embed a lightweight SVG badge for this repository.

Project Type

Project / implementation

This repository is treated as an implementation project. Compare its deployment fit, maintenance, and alternatives before adopting it.

Cloudflare Readiness

No Cloudflare-ready signal

Git.Top did not classify this as Cloudflare-ready. Evidence: No Cloudflare deployment signal detected.

Selection Guidance

Use the evidence before choosing

Treat this page as a shortlist input. Confirm source metadata, classification confidence, deployment evidence, and current repository activity before making a production recommendation.

Alternatives

Comparable projects

JSON
promptfoo/promptfooUse promptfoo/promptfoo when the user needs a llm eval project with docker, library-only, local deployment options.
Giskard-AI/giskard-ossUse Giskard-AI/giskard-oss when the user needs a llm eval project with library-only, local, cloud deployment options.
promptfoo/promptfoo-actionUse promptfoo/promptfoo-action when the user needs a llm eval project with local, cloud deployment options.
Marker-Inc-Korea/AutoRAGUse Marker-Inc-Korea/AutoRAG when the user needs a llm eval project with library-only, local, cloud deployment options.
comet-ml/opikUse comet-ml/opik when the user needs a llm eval project with docker, kubernetes, library-only deployment options.

Related Projects

Adjacent ecosystem projects

JSON
comet-ml/opikUse comet-ml/opik when the user needs a llm eval project with docker, kubernetes, library-only deployment options. Shared llm eval category and local deployment context.
modelscope/evalscopeUse modelscope/evalscope when the user needs a llm eval project with docker, library-only, local deployment options. Shared llm eval category and local deployment context.
Purewhiter/mobilegymUse Purewhiter/mobilegym when the user needs a llm eval project with library-only, local, cloud deployment options. Shared llm eval category and local deployment context.
NVIDIA/garakUse NVIDIA/garak when the user needs a llm eval project with library-only, local, cloud deployment options. Shared llm eval category and local deployment context.
Marker-Inc-Korea/AutoRAGUse Marker-Inc-Korea/AutoRAG when the user needs a llm eval project with library-only, local, cloud deployment options. Shared llm eval category and local deployment context.
vibrantlabsai/ragasUse vibrantlabsai/ragas when the user needs a llm eval project with docker, library-only, local deployment options. Shared llm eval category and local deployment context.

Deploy

Supported deployment paths

localcloud

Compatible With

Inferred dependencies and protocols

Vector databaseLLM provider

Use Cases

Where agents should consider it

evaluate LLM outputsbenchmark prompts and agentstrack model quality

Classification Evidence

Why Git.Top categorized this project

Full JSON
Category: highMatched "eval" in metadata. Matched "benchmark" in metadata.
Deployment: mediumLocal usage is assumed for open source repositories unless contradicted.
Difficulty: mediumRepository has under 10k stars, so complexity is treated conservatively.
Cloudflare Ready: highNo Cloudflare deployment signal detected. No wrangler.toml found in inspected repository files.

Compare

promptfoo/promptfoo leads this context

JSON
ProjectStarsAgentLocalCloudflareScore
supermemoryai/memorybench303NoYesNo55
promptfoo/promptfoo23,937NoYesNo88
Giskard-AI/giskard-oss5,729YesYesNo81
promptfoo/promptfoo-action69NoYesNo58