Skip to main content

Ponytail Gain

Display this scoreboard when invoked. One-shot: do NOT change mode, write flag files, or persist anything.

The figures are the published agentic benchmark averages: a headless Claude Code session (Haiku 4.5) doing 12 feature tickets on a real FastAPI + React repo, each task averaged over 4 runs, against the same agent without the skill, plus 6 safety tasks. They are measured, not computed from the current repo. Source: benchmarks/results/2026-06-18-agentic.md and the README.

Scoreboard​

Render plain ASCII bars. The bar length shows ponytail as a share of the no-skill baseline; the label carries the exact figure:

ponytail gain benchmark average · 12 tasks · Haiku 4.5

no-skill ████████████████████ 100%
Lines of code █████████··········· 46% ▼ 54%
Tokens ████████████████···· 78% ▼ 22%
Cost ████████████████···· 80% ▼ 20%
Time ███████████████····· 73% ▼ 27%
Safety kept 100% (validation, error handling, security)

This repo: /ponytail-debt (shortcuts you deferred)
/ponytail-audit (what's still cuttable)

Honesty boundary​

These are benchmark averages, not this repo. NEVER print a per-repo savings number ("you saved X lines/tokens here"): the unbuilt version was never written, so there is no real baseline to subtract from in a live repo. The only real per-repo figures come from /ponytail-debt (a counted ledger), and this card points there instead of inventing one.

Boundaries​

One-shot display. Edits nothing, changes no mode. "stop ponytail" or "normal mode": revert.


Source: jknash/hermes-shared-skills · branch hermes-jkdev001 @ 1d0d545c3970 · skills/software-development/ponytail-gain/ · view source · Imported 2026-10-04. Supporting files (references, scripts) remain in the source repository.

Published by Muse · 2026-10-04.

Updated 2026-10-07 to upstream DietrichGebert/ponytail v4.13.0 (08e952d7a8) — body replaced with the upstream text; no hooks, scripts, or plugins imported.