# Benchmark agent memory and RAG systems with MemoryBench

> Use MemoryBench to run repeatable conversational memory and RAG benchmarks across providers, datasets, judge models, checkpoints, and structured reports.

- Skill: `agentskillexchange/benchmark-agent-memory-and-rag-systems-with-memorybench` (Agent Skill)
- Install (CLI): `npx skillmds@latest add agentskillexchange/benchmark-agent-memory-and-rag-systems-with-memorybench`
- Raw SKILL.md: https://api.skillmd.com/api/skills/agentskillexchange/benchmark-agent-memory-and-rag-systems-with-memorybench/raw
- Safety review: pending (external: skill-scanner PASS, skillspector PASS)
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: AI & ML
- Author: agentskillexchange (https://skillmd.com/u/agentskillexchange)
- Updated: 2026-09-08
- Page: https://skillmd.com/skills/agentskillexchange/benchmark-agent-memory-and-rag-systems-with-memorybench

---


# Benchmark agent memory and RAG systems with MemoryBench

Use MemoryBench to run repeatable conversational memory and RAG benchmarks across providers, datasets, judge models, checkpoints, and structured reports.

## Prerequisites

Bun, MemoryBench repository, at least one memory/RAG provider API key, at least one judge model API key, benchmark datasets

## Installation

Basic usage or getting-started notes:
- 🆚 Multi‑provider comparison: run the same benchmark across providers side‑by‑side
- 📊 Structured reports: export run status, failures, and metrics for analysis
- bun install

- Source: https://github.com/supermemoryai/memorybench
- Extracted from upstream docs: https://raw.githubusercontent.com/supermemoryai/memorybench/HEAD/README.md

## Documentation

- https://supermemory.ai/docs/memorybench/overview

## Source

- [Agent Skill Exchange](https://agentskillexchange.com/skills/benchmark-agent-memory-and-rag-systems-with-memorybench/)

