Source profileQuality 94/100Review permissions

MichelKerkmeester/opencode--skilled-agent-loops-with-spec-kit-memory/.opencode/skills/sk-doc/sk-create-benchmark/SKILL.md

sk-create-benchmark

Author MCP-promotion, behavior, skill-benchmark, and model-benchmark artifacts; route the Lane A authoring guide.

Source repository stars
34
Declared platforms
0
Static risk flags
2
Last source update
2026-08-28
Source checked
2026-08-28

Decision brief

What it does: where it fits

create-benchmark is the sk-doc benchmark-authoring packet. It covers:

Best for

  • Activation Triggers
  • Adoption Gate (MCP promotion)
  • Trigger Signals

Not for

  • Tasks that require unconfirmed production actions or broad system permissions.
  • Environments where the pinned source and install steps cannot be inspected.

Compatibility matrix

Platform support, with evidence labels

PlatformStatusEvidenceWhat to check
CodexNot declaredNo explicit evidencePortability before use
Claude CodeNot declaredNo explicit evidencePortability before use
CursorNot declaredNo explicit evidencePortability before use
Gemini CLINot declaredNo explicit evidencePortability before use
Open the compatibility checker

Installation

Inspect first. Install second.

The source command is displayed only when detected. A safe inspection prompt is always available so your agent can explain every action before execution.

Source-detected install commandSource
npx skills add https://github.com/MichelKerkmeester/opencode--skilled-agent-loops-with-spec-kit-memory --skill ".opencode/skills/sk-doc/sk-create-benchmark"
Safe inspection promptEditorial

Inspect the Agent Skill "sk-create-benchmark" from https://github.com/MichelKerkmeester/opencode--skilled-agent-loops-with-spec-kit-memory/blob/3d386ee21366523774d89c0aff3ebbbc8fa7ff10/.opencode/skills/sk-doc/sk-create-benchmark/SKILL.md at commit 3d386ee21366523774d89c0aff3ebbbc8fa7ff10. List every install step, command, network request, credential, file read/write, external action, and rollback step. Explain whether it fits my task. Do not install or execute anything until I approve.

Workflow

What the source asks the agent to do

  1. 01

    4. HOW IT WORKS: AUTHORING WORKFLOW

    Complete these steps in order after the spec packet ships.

    Confirm the promotion gate. Read decision-record.md, implementation-summary.md, and benchmark evidence. Require an accepted decision, stable headline and fixture, replay commands, and a defensible winner or explicit pro…Classify the task. Decide whether this is a true promotion, a re-run update, or a retirement update.Confirm the target skill. Verify the consuming skill has mcp-server/ and an appropriate measurable MCP surface.
  2. 02

    Templates, Workflow, and Naming

    Load the behavior-benchmark guide and shared framework before authoring. The guide owns templates, sequence, matrix, and naming; execution and evidence stay in the executing packet.

    Load the behavior-benchmark guide and shared framework before authoring. The guide owns templates, sequence, matrix, and naming; execution and evidence stay in the executing packet.
  3. 03

    Authoring Workflow

    1. Read the storage guide — confirm run-label naming and frozen baseline. 2. Confirm the target has (or is establishing) a Lane C benchmark/ tree beside the skill it measures. 3. Author the index from the template: newest-first folder rows, structure map, re-run commands, and li…

    Read the storage guide — confirm run-label naming and frozen baseline.Confirm the target has (or is establishing) a Lane C benchmark/ tree beside the skill it measures.Author the index from the template: newest-first folder rows, structure map, re-run commands, and links to scoring and /deep:skill-benchmark.
  4. 04

    1. WHEN TO USE

    Use this packet to author completed benchmark evidence or benchmark inputs into the skill tree. Route through §2 first; families are distinct.

    MCP promotion (on-disk shared; §3-8) — promote a completed MCP benchmark from a spec packet: author the ten-section benchmark-report.md and source.md, copy results.csv, applicable per-probe.jsonl and runtime sidecars in…Behavior benchmark (§9) — author or extend a deep-loop mode's index, -NNN-.md scenario contracts, baseline, and entry-surface/clarity matrix. Fixed prefixes are research (RSB), review (RVB), ai-council (ACB), and improv…Skill-benchmark (§10) — establish Lane C sibling run-label folders with frozen baseline/, or author/update benchmark/README.md.
  5. 05

    Activation Triggers

    Keyword triggers: benchmark-report.md, source.md, mcp-server/benchmarks, MCP bake-off; behavior benchmark, behavior-benchmark.md, behaviorbenchmark, scenario contract, benchmark/README.md, run-label folder, benchmark package; model-benchmark, benchmark fixture, benchmark profile…

    MCP promotion (on-disk shared; §3-8) — promote a completed MCP benchmark from a spec packet: author the ten-section benchmark-report.md and source.md, copy results.csv, applicable per-probe.jsonl and runtime sidecars in…Behavior benchmark (§9) — author or extend a deep-loop mode's index, -NNN-.md scenario contracts, baseline, and entry-surface/clarity matrix. Fixed prefixes are research (RSB), review (RVB), ai-council (ACB), and improv…Skill-benchmark (§10) — establish Lane C sibling run-label folders with frozen baseline/, or author/update benchmark/README.md.

Permission review

Static risk signals and limitations

Writes files

medium · line 35

The documentation asks the agent to create, modify, or delete local files.

Create a skill-local MCP-promotion folder only when all apply:

Writes files

medium · line 48

The documentation asks the agent to create, modify, or delete local files.

YES -> Create a benchmark folder

Runs scripts

medium · line 186

The documentation asks the agent to run terminal commands or scripts.

python3 .opencode/skills/sk-doc/shared/scripts/check_authored_name_kebab.py <artifact-path-or-slug>

Evidence record

Why each signal appears

EvidenceSourceComputedTestedEditorial
SignalValueEvidence typeMeaning
Quality score94/100ComputedDocumentation, specificity, maintenance, and trust rules
Repository stars34SourceRepository attention, not individual Skill quality
Compatibility0 platformsSourceDeclared in the catalog source record
Usage guideautomated source guideEditorialGenerated or reviewed according to the visible evidence level

Pinned source

Provenance and original SKILL.md

Repository
MichelKerkmeester/opencode--skilled-agent-loops-with-spec-kit-memory
Skill path
.opencode/skills/sk-doc/sk-create-benchmark/SKILL.md
Commit
3d386ee21366523774d89c0aff3ebbbc8fa7ff10
License
MIT
Collected
2026-08-28
Default branch
main
View the original SKILL.md

Frequently asked questions

What to verify before installation and use

What does the sk-create-benchmark source document cover?

create-benchmark is the sk-doc benchmark-authoring packet. It covers:

How do I install sk-create-benchmark?

The source record exposes this install command: npx skills add https://github.com/MichelKerkmeester/opencode--skilled-agent-loops-with-spec-kit-memory --skill ".opencode/skills/sk-doc/sk-create-benchmark". Inspect the command and pinned source before running it.

Which permission-related actions were detected?

Static rules flagged write-files, exec-script in the source; the page lists the matching lines and excerpts.

Alternatives

Compare before choosing

Computed 10045,960

coreyhaines31/marketingskills

ab-testing

When the user wants to plan, design, or implement an A/B test or experiment, or build a growth experimentation program. Also use when the user mentions "A/B test," "split test," "experiment," "test this change," "variant copy," "multivariate test," "hypothesis," "should I test this," "which version is better," "test two versions," "statistical significance," "how long should I run this test," "growth experiments," "experiment velocity," "experiment backlog," "ICE score," "experimentation program

Computed 10045,960

coreyhaines31/marketingskills

churn-prevention

When the user wants to reduce churn, build cancellation flows, set up save offers, recover failed payments, or implement retention strategies. Also use when the user mentions 'churn,' 'cancel flow,' 'offboarding,' 'save offer,' 'dunning,' 'failed payment recovery,' 'win-back,' 'retention,' 'exit survey,' 'pause subscription,' 'involuntary churn,' 'people keep canceling,' 'churn rate is too high,' 'how do I keep users,' or 'customers are leaving.' Use this whenever someone is losing subscribers o

Computed 10025,136

alirezarezvani/claude-skills

app-store-optimization

App Store Optimization (ASO) toolkit for researching keywords, analyzing competitor rankings, generating metadata suggestions, and improving app visibility on Apple App Store and Google Play Store. Use when the user asks about ASO, app store rankings, app metadata, app titles and descriptions, app store listings, app visibility, or mobile app marketing on iOS or Android. Supports keyword research and scoring, competitor keyword analysis, metadata optimization, A/B test planning, launch checklist

Computed 10015,385

wanshuiyin/Auto-claude-code-research-in-sleep

citation-audit

Use it for operations and research tasks; the detail page covers purpose, installation, and practical steps.