SkillAgentSearch skills...

gentle-ai-bench

Trigger: bench, journey, journeys, driven mode, gentle-ai-bench, journey corpus, j-numbers, bench axis. Author and verify gentle-ai bench journeys; go test ./bench never proves driven execution.

Install / Use

npx skills add Gentleman-Programming/gentle-ai --skill gentle-ai-bench

Installs into whichever agent you are using.

About this skill
📄

SKILL.md

Installable skill definition

Quality Score

80/100

Category

Legal

Supported Platforms

Universal

Tags

Our assessment of gentle-ai-bench

gentle-ai-bench scores 80/100 on our quality scale, 77th of 103 Legal skills we index.

Its SKILL.md is 3.1 KB long, split into 4 sections and no code examples: a solid amount of guidance for an agent.

With 7,305 GitHub stars, it is one of the more widely adopted skills in the catalogue.

Substance
26/30
Structure
8/20
Description
15/15
Adoption
16/20
Freshness
15/15

Maintenance, license and trust

  • The repository was last updated 2 days ago, so gentle-ai-bench is actively maintained.
  • It is released under the MIT license, a permissive license that allows use, modification and commercial use with attribution.
  • Its trust signals score 100/100, with no cautions. These come from repository metadata, not a code audit — read the skill file before letting an agent act on it.

gentle-ai-bench compared with similar skills

All 4 of these similar skills score higher than gentle-ai-bench; compare them before choosing.

SkillScoreStarsUpdatedFormat
gentle-ai-bench (this skill)by Gentleman-Programming807.3k2d agoSKILL.md
algorithmic-artby anthropics100177.9k5d agoSKILL.md
pptxby anthropics100177.9k5d agoSKILL.md
designby nextlevelbuilder100130.2k6d agoSKILL.md
ui-ux-pro-maxby nextlevelbuilder100130.2k6d agoSKILL.md

Frequently asked questions

How do I install gentle-ai-bench?
Run npx skills add Gentleman-Programming/gentle-ai --skill gentle-ai-bench. The install tabs above show the steps for each supported agent.
Which AI agents does gentle-ai-bench work with?
It is written for Universal, as a SKILL.md file. Other agents that read the same format can often use it too.
Is gentle-ai-bench safe to use?
It is MIT-licensed and scores 100/100 on trust signals. Skills are instructions an agent will follow, so read the file before installing it and do not approve commands you do not understand.
Is gentle-ai-bench still maintained?
The repository was last updated 2 days ago, so gentle-ai-bench is actively maintained.

name: gentle-ai-bench description: "Trigger: bench, journey, journeys, driven mode, gentle-ai-bench, journey corpus, j-numbers, bench axis. Author and verify gentle-ai bench journeys; go test ./bench never proves driven execution." license: Apache-2.0 metadata: author: "Gentleman-Programming" version: "1.0"

Activation Contract

Load when touching bench/ in gentle-ai, adding or changing a journey, changing a product semantic a journey might pin, or diagnosing a bench failure in CI's Unit Tests job.

Hard Rules

  • go test ./bench validates corpus declarations only. It does NOT execute journeys. The only driven proof is building the harness and the product binary and running the harness against it; a green go test ./bench claims nothing about execution.
  • Reproduce CI, do not guess invocations: read the Unit Tests step in .github/workflows/ci.yml and copy its exact build and gentle-ai-bench run --binary ... commands. Use --only <journey-id> to drive one journey.
  • Journey IDs are unique across every journeys_*.go file. The collision guard fails loudly naming both files; pick an unused ID by reading the corpus, never reuse a retired one.
  • Every journey declares Review: — reviewOptedIn (the runner enables receipt-driven development globally before the first step, uncounted, and fails the journey if the switch does not come on) or reviewUntouched (its subject IS the switch, or it has nothing to do with reviews). The declaration is mandatory; validateCorpus fails the run without it. Lifecycle journeys must not depend on the product default. Reviews default to ON; reviewUntouched does not imply OFF. Journeys requiring OFF must explicitly disable it, while default-mode journeys must assert ON/default with unset sources.
  • Every execute transition must carry a runnable command; the dead-execute guard fails the run otherwise.
  • When a ratified product semantic changes, grep the corpus for journeys pinning the OLD behavior before shipping. The corpus is a second test surface beyond unit tests; a journey asserting the defect keeps the defect green.
  • dead_end prints n/a unless the run actually measured one. Never fabricate a value to move the column.
  • A by_design exemption costs a shape from the closed vocabulary plus a verified quote of the product's own next-action text. If the quote no longer tells the operator what to do, it is a defect wearing an exemption.
  • Prefer a NEW journeys_*.go file when the shared ones are owned by open PRs; bump the core journey-count pin in the same change.

Execution Steps

  1. Read the corpus area you touch and the CI invocation before writing.
  2. Author or adapt the journey; update its title, step names, and comment to say WHY the expectation holds (cite the issue or ratified decision).
  3. Run go test ./... in bench/ for declarations, THEN the driven harness for execution; both results go in the PR body.
  4. On semantic changes, list the journeys you checked for stale pins.

Output Contract

PR evidence includes the driven-mode summary line (completed / unsupported / failed counts) from a locally built binary, not only go test output.

Related Skills

View on GitHub
GitHub Stars7.3k
CategoryLegal
Updated2d ago
Forks799

Languages

Go

Trust signals

100/100

From repository metadata: license, adoption, age and documentation. Not a code audit — see the Safety scan above for what the skill file itself contains.

No cautions