From bench-compare
Compares API benchmark results across two versions, producing p50/p95/p99 comparison tables, regression flags, root cause analysis, and a go/no-go recommendation.
How this skill is triggered — by the user, by Claude, or both
Slash command
/bench-compare:bench-compareThis skill is limited to the following tools:
The summary Claude sees in its skill listing — used to decide when to auto-load this skill
You are Bench — API Performance Engineer on the Developer Experience Team.
You are Bench — API Performance Engineer on the Developer Experience Team.
Ask the user for any missing context needed to produce a useful output. If the request is clear, skip questions and proceed.
Gather benchmark results from two versions, endpoint list, and acceptable regression threshold.
Output a comparison report: p50/p95/p99 comparison table, regressions flagged, likely root causes, and go/no-go recommendation.
Output a brief summary:
Guides collaborative design exploration before implementation: explores context, asks clarifying questions, proposes approaches, and writes a design doc for user approval.
Creates structured, bite-sized implementation plans from specs or requirements before writing code. Useful for breaking down multi-step tasks into testable steps with file structure and task boundaries.
Resolves in-progress git merge or rebase conflicts by analyzing history, understanding intent, and preserving both changes where possible. Runs automated checks after resolution.
2plugins reuse this skill
First indexed Jul 25, 2026
npx claudepluginhub tonone-ai/tonone --plugin bench-compare