Project comparison
Compare adoption, momentum, maintenance health, and project basics before choosing which tool to evaluate deeper.
Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI and Anthropic.
Best matched with other agent tools.
Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with Claude Code, OpenClaw, Codex, Gemini, Hermes, Copilot, OpenCode + More
Best matched with other agent tools.
claude-mem has the larger GitHub footprint with 88.2K stars.
claude-mem is currently growing faster at +13.4K stars this week.
claude-mem has the stronger automated maintenance signal at 98/100. This is not a security or fit verdict.
Use these signals to narrow your choice, then confirm setup, license, and fit upstream.
| Signal | Promptfoo | claude-mem |
|---|---|---|
| Evidence status | Basic listing· 2d ago | Recently verified· 2d ago |
| GitHub stars | 23.5K | 88.2K |
| Weekly growth | +2.3K | +13.4K |
| Health score | Strong93/100 |
Get the fastest-growing projects, useful MCP servers, and technical reads in one weekly email.
| Contributors | 320 | 122 |
|---|
| Commits per week | 74.7 | 30.1 |
|---|
| Open issues | 437 | 354 |
|---|
| Language | TypeScript | JavaScript |
|---|
| License | MIT | Apache-2.0 |
|---|
| Last commit | 2d ago | 4d ago |
|---|
| Last release | 0.121.19 | v13.11.0 |
|---|