- name
- zouroboros-bench
- description
- Benchmark harness for AI memory systems. Evaluates LongMemEval, LoCoMo, and ConvoMem datasets against any memory backend via the zouroboros-memory CLI. Includes Mimir judge for catching architectural drift.
- version
- 1.0.0
- compatibility
- OpenClaw, Claude Code, Codex CLI, any Node.js 22+ environment
- metadata
- author
- marlandoj.zo.computer
- openclaw
- emoji
- 📊
- requires
- bins
- [node]
- env
- [OPENAI_API_KEY]
- primaryEnv
- OPENAI_API_KEY
- install
- kind
- node
- package
- zouroboros-bench
- bins
- [zouroboros-bench, zouroboros-bench-report]
- label
- Install Zouroboros Bench (npm)
- homepage
- https://github.com/AlaricHQ/zouroboros-openclaw
Usage
Install: npm install zouroboros-bench zouroboros-memory
Run all benchmarks
npx zouroboros-bench --limit 50Run specific benchmark
npx zouroboros-bench --benchmarks longmemeval --limit 100 --judgeGenerate report
npx zouroboros-bench-report --runs ./data/runs/Environment Variables
ZOUROBOROS_MEMORY_CLI— Path to memory CLI binary (default:zouroboros-memory)ZOUROBOROS_MEMORY_DB— SQLite DB path for benchmarksOPENAI_API_KEY— Required for GPT-4o judgeOLLAMA_URL— Ollama URL for local LLM (default:http://localhost:11434)