Benchmark agent memory and RAG systems with MemoryBench

Use MemoryBench to run repeatable conversational memory and RAG benchmarks across providers, datasets, judge models, checkpoints, and structured reports.

agentskillexchange Updated 28 repo stars

File contents

Benchmark agent memory and RAG systems with MemoryBench

Use MemoryBench to run repeatable conversational memory and RAG benchmarks across providers, datasets, judge models, checkpoints, and structured reports.

Prerequisites

Bun, MemoryBench repository, at least one memory/RAG provider API key, at least one judge model API key, benchmark datasets

Installation

Basic usage or getting-started notes:

Documentation

Source

agentskillexchange/skills/tree/main/skills/benchmark-agent-memory-and-rag-systems-with-memorybench commit dec68c1630

Frequently asked questions

npx skillmds@latest add agentskillexchange/benchmark-agent-memory-and-rag-systems-with-memorybench