Claude Skill

Benchmark agent memory and RAG systems with MemoryBench

Use MemoryBench to run repeatable conversational memory and RAG benchmarks across providers, datasets, judge models, checkpoints, and structured reports.

LLM Mart · 0 points · 0 views 0 listing impressions 0 install-command copies
Virus-scanned Reviewed automatically before listing.

Full trust report

Download agentskillexchange-skills-skills_benchmark-agent-memory-and-rag-systems-with-memorybench-07beb56.zip · 0 KB
Part of agentskillexchange/skills — 249 skills

Install

skills CLI npx skills add https://github.com/agentskillexchange/skills/tree/main/skills/benchmark-agent-memory-and-rag-systems-with-memorybench
Claude Code claude plugin marketplace add https://llmmart.ai/marketplace.json && claude plugin install agentskillexchange-skills@llmmart
Git git clone https://github.com/agentskillexchange/skills.git

The skills CLI installs just this skill, for any of its supported agents. Claude Code installs the whole agentskillexchange/skills collection as a plugin from our marketplace. Git is the plain clone.

Skill manifest

Benchmark agent memory and RAG systems with MemoryBench

Use MemoryBench to run repeatable conversational memory and RAG benchmarks across providers, datasets, judge models, checkpoints, and structured reports.

Prerequisites

Bun, MemoryBench repository, at least one memory/RAG provider API key, at least one judge model API key, benchmark datasets

Installation

Basic usage or getting-started notes:

Documentation

Source

Files (skills)
  • SKILL.md 1.5 KB
    ---
    name: "Benchmark agent memory and RAG systems with MemoryBench"
    slug: "benchmark-agent-memory-and-rag-systems-with-memorybench"
    description: "Use MemoryBench to run repeatable conversational memory and RAG benchmarks across providers, datasets, judge models, checkpoints, and structured reports."
    github_stars: 293
    verification: "listed"
    source: "https://github.com/supermemoryai/memorybench"
    author: "Supermemory contributors"
    publisher_type: "open_source"
    category: "Runbooks & Diagnostics"
    framework: "Multi-Framework"
    tool_ecosystem:
      github_repo: "supermemoryai/memorybench"
      github_stars: 293
    ---
    
    # Benchmark agent memory and RAG systems with MemoryBench
    
    Use MemoryBench to run repeatable conversational memory and RAG benchmarks across providers, datasets, judge models, checkpoints, and structured reports.
    
    ## Prerequisites
    
    Bun, MemoryBench repository, at least one memory/RAG provider API key, at least one judge model API key, benchmark datasets
    
    ## Installation
    
    Basic usage or getting-started notes:
    - 🆚 Multi‑provider comparison: run the same benchmark across providers side‑by‑side
    - 📊 Structured reports: export run status, failures, and metrics for analysis
    - bun install
    
    - Source: https://github.com/supermemoryai/memorybench
    - Extracted from upstream docs: https://raw.githubusercontent.com/supermemoryai/memorybench/HEAD/README.md
    
    ## Documentation
    
    - https://supermemory.ai/docs/memorybench/overview
    
    ## Source
    
    - [Agent Skill Exchange](https://agentskillexchange.com/skills/benchmark-agent-memory-and-rag-systems-with-memorybench/)
    

Comments (0)

Sign in to join the conversation.

No comments yet.

Reviews (0)

No reviews yet.

Related