Third-party skills gap analysis for the agent fleet
Date: 2026-10-07 · Access date for all citations: 2026-10-07
Prepared by the fleet research service: planned and verified by the lead researcher (maverick-muse-lead_researcher-001); executed by the research assistant (maverick-muse-research_assistant-001) using deep research against the live docsite index, the upstream repositories, and the public skills registries named below. Read-only throughout: nothing was installed, imported, or executed.
Question. The owner: "We have a lot of skills installed and I want to make sure we're using them effectively. I want to utilize some of the third-party skills like no-ai-slop and ponytail. Dispatch my research team to identify any skills we don't have available, but should, and get them into docsite and the agent runbook configurations." A mid-mission amendment from the owner added agile project management skills, aimed mainly at the scrum master (§6, rows 16–28).
Boundary. This report delivers the vetted list only. Import into the docsite and the runbook configurations happens after the chief's (and where needed the owner's) approval of the list; the chief routes that work to the sweeper. The research service imports nothing.
1. Bottom line
The owner-named skills are already in the library, the assignment's starting sweep was wrong, as the lead's intake correction suspected and this mission verified from the live index: skill-no-ai-slop and a complete six-page ponytail family are in docsite's 429-skill index. Their real gaps sit in completeness, currency, and wiring: the no-ai-slop page's process references a companion eval.md that was never imported; the ponytail import is an untracked snapshot (upstream DietrichGebert/ponytail is at v4.13.0 (2026-10-05) and its recent releases have already changed /ponytail-gain's basis); the sibling writing skill humanizer is materially stale (fleet 2.5.1 vs upstream 3.1.0); and none of the sampled non-research role runbooks name any skills at all, no-ai-slop and ponytail sit in manifests that the chief/worker/reviewer runbooks never reference. Beyond the named families, the true acquisition gaps are few and specific: Anthropic's official frontend-design (absent; Apache-2.0, no scripts), a dedicated WCAG/accessibility audit method (absent, AccessLint/skills is the credible candidate, with license and execution caveats), and the systemic ui_ux coverage hole (49 creative-category pages, zero ui_ux tags, no ui_ux manifest). Superpowers (15/15) and the anthropics/skills set (17/19 in some form) are already essentially fully held, the fleet's problem there is duplication (eight skills are held twice, under both the anthropic and the claude categories) and wiring, not acquisition.
Agile (Amendment 1). The owner added agile project skills to the mission mid-flight, aimed at the scrum master. The internal scrum set is an operations set of 18 skills (retro, handoffs, to-tickets, triage, status reporting), and the RUNBOOKS.md pointer to a "planning & agile method set" names a set that does not exist as a set in the docsite. Against the fleet's settled method (sprints close on scope completion, never on a timebox; the in-repo llm-kanban board; Fibonacci points; two-stage intake; owner acceptance), exactly two external candidates clear the bar, both adapted: plinth's agile story trio and three plays from tronghieu's scrum-master set, plus a zero-cost wire-in (planning-and-task-breakdown into the scrum manifest, where the chief's disposition and the manifest on main currently disagree). Section 6 gives the per-sub-area evidence; rows 16–28 of the acquisition list carry the calls.
2. Baseline inventory + capability coverage map
Baseline (verified live, docsite main): skills index docs/15-skills/skills-index.json blob 5f5777e794, 429 entries, the lead's intake numbers reproduce exactly. Index role tags: worker 185, chief_of_staff 126, planner 82, code_quality_reviewer 44, sweeper 36, security_reviewer 26, research_assistant 18, scrum_master 15, lead_researcher 8, finance 2, all 3. Per-role manifests on main (10 files) run +3 vs index tags (the three all skills): worker 188, chief 129, planner 85, CQR 47, sweeper 39, security_reviewer 29, research_assistant 21, scrum_master 18, lead_researcher 11, finance 5. No ui_ux manifest exists and the string ui_ux appears in zero entries' roles (index query: any role == ui_ux → 0 hits). Workspace skills on this host (ls ~/workspace/skills): exactly the five named, cloudflare, cloudflare-dedicated, fleet-sync, github-docsite, llm-kanban. Bundled platform catalog appendix: ~50 product skills (gmail, calendar, spotify, shopping, plaid, etc.) ride with the platform, a different layer from the docsite method library; full enumeration in my working notes on request, none are fleet method skills, all out of this mission's scope.
Coverage map (index-derived; method stated so counts are auditable):
| Area | Evidence from the index | Verdict |
|---|---|---|
| Writing quality / anti-slop | skill-no-ai-slop + skill-humanizer, both tagged chief_of_staff only; 33 purpose-text hits for writing/editing terms. Division of ground: humanizer = pattern catalogue for removing AI tells (Wikipedia-derived, 34 patterns as imported); no-ai-slop = an editor's process (preserve voice, minimum edit, detect-vs-edit modes, self-check vs eval.md). They overlap in aim, differ in method, but only the chief is tagged for either; worker/planner/research roles that also write have no tagged anti-slop skill | Present, narrowly tagged, one stale |
| Coding discipline / minimalism | Ponytail family (6) + skill-clean-code, skill-code-simplification, skill-simplify-code, skill-coding-behavior-rules, skill-lean-build, skill-minimalist-skill | Strong |
| Code review | 16 ids containing "review" (incl. skill-code-review, skill-code-review-and-quality, skill-codebase-review, superpowers requesting/receiving) | Strong |
| Planning | skill-writing-plans, skill-executing-plans, skill-planning-and-task-breakdown, skill-plan; planner has 82 tags | Strong |
| Design / UI-UX | creative category = 49 pages; 17 ids containing "design"; strong exemplars: skill-ui-ux-pro-max, skill-design, skill-design-system, skill-emil-design-eng, skill-impeccable, skill-redesign-skill. BUT roles are worker/CQR/planner, ui_ux: 0 tags, no manifest, and (see §5) the ui_ux runbook names no skills | Content strong, role coverage ZERO, the map's biggest hole |
| Accessibility | 0 ids matching accessib*/wcag*; "accessibility" appears in the purpose text of only 5 entries (frontend-ui-engineering, react-patterns, react-testing, ui-styling, ui-ux-pro-max, incidental, incl. ui-ux-pro-max's searchable data). No WCAG audit method anywhere | GAP, acquire |
| Testing | 13 ids containing "test" (TDD ×2 incl. an Osmani variant, webapp-testing ×2 duplicate, e2e, browser-devtools, language-specific) | Adequate; duplicative in places |
| Security review | 5 ids containing "secur" + catalog-driven security testing, security-and-hardening; security_reviewer 26 tags | Adequate |
| Research method | research category 16 pages; 15 ids containing "research"; research_assistant manifest 21, lead 11 | Strong for research roles |
| Memory / context | 8 ids containing "context" + skill-memory-systems, skill-codebase-memory, context-engineering family, mostly chief-tagged | Adequate |
| Agile project method (Amendment 1) | Scrum manifest = 18 operations skills; vocabulary probe of all 429 entries: "sprint" 0, "groom" 0, "standup" 0, "velocity" 0, "agile" 0, "user stor" 0; planning-and-task-breakdown tagged [planner] only | Thin by design in places (estimation settled by charter; DoR/DoD enforced by the llm-kanban schema); genuine gaps in story craft and impediment handling, see §6 |
| Documentation quality | doc-coauthoring held twice (anthropic + claude duplicates), skill-documentation-and-adrs, changelog-generator; no dedicated technical-writing-quality skill beyond coauthoring | Adequate, duplicative; thin at the quality end |
Duplication datum: docs/15-skills/anthropic/ and docs/15-skills/claude/ are the same 8 skills imported twice (algorithmic-art, canvas-design, claude-api, doc-coauthoring, internal-comms, mcp-builder, web-artifacts-builder, webapp-testing, anthropic-* and claude-* ids, verified by index lookup). The claude category holds three further skills (code, design, handoff) that are not duplicates. ~8 of 429 index slots are double-counts.
3. Named-skills dossiers
3a. no-ai-slop, HELD; incomplete import; roles too narrow
- Docsite page (fetched in full):
docs/15-skills/creative/no-ai-slop.md, blob 3b63bfa40d, 11,307 chars, a full skill text (edit + detect modes, editing principles, pattern list, process), NOT a summary. Footer attribution: sourcejknash/hermes-shared-skillsbranch hermes-jkdev001 @1d0d545c3970, imported 2026-10-04; "Supporting files (references, scripts) remain in the source repository." - Completeness gap (verified on the page): its Process step 4 says "check the edited draft against
eval.md", that companion file is NOT in docsite (the footer confirms supporting files stayed behind). The self-check half of the method is unusable as imported. - True upstream (verified at repo): https://github.com/petergyang/no-ai-slop, Peter Yang, MIT License, ~12,004★, created 2026-07-07, 7 releases to v1.0.6; the skill lives at
skills/no-ai-slop/{SKILL.md, eval.md, agents/}upstream, i.e. upstream ships exactly the eval.md the docsite copy lacks. Repo also carries scripts/, .codex-plugin, build_plugin.py (build/validate tooling), PRIVACY.md/TERMS.md, none of that needs importing; the skill + eval.md do. Whether the docsite text matches upstream v1.0.6 line-for-line was NOT diffed (gap, §8). - Role coverage: tagged chief_of_staff only (index + chief manifest). Plausible users who are untagged: planner, lead_researcher/research_assistant (long-form reports), worker (PR prose). Runbook wiring: none (§5).
- Confidence: Settled (presence, page completeness, eval.md absence, upstream identity/license/version); Supported (role-extension judgement).
3b. Ponytail family, HELD, family COMPLETE vs upstream; currency untracked; executable ecosystem correctly NOT imported
- Docsite pages (all fetched in full):
skill-ponytail(6,196 chars, blob 8e794238ac), ponytail-review (2,398), ponytail-audit (1,734), ponytail-debt (1,807), ponytail-gain (1,989), ponytail-help (3,058), full skill texts with the same internal-mirror footer (hermes-shared-skills @1d0d545c3970, imported 2026-10-04); ponytail main page carries "license MIT.". Roles: ponytail/gain/help → worker; review/audit → code_quality_reviewer; debt → CQR + sweeper. - Upstream (verified at repo): https://github.com/DietrichGebert/ponytail, Dietrich Gebert, MIT License (LICENSE file + sidebar), ~157k★, 311 commits, 22 releases, latest v4.13.0 (released 2026-10-05), repo created 2026-06-12, active Oct 2026. Upstream
skills/= exactly the same 6 + plugin.json, the fleet's family is complete; upstream has no seventh skill. - Currency: upstream moves fast and has changed around the skills: recent releases number review/audit findings and changes /ponytail-gain to agentic benchmarks, the fleet's gain page still describes the older published-medians scoreboard, so at least that page is behind upstream. The docsite import is a snapshot of an internal mirror, not tracked to any upstream release; exact per-page diff vs the current upstream release not performed (gap, §8).
- Trust note (important): upstream is no longer markdown-only, it ships JS hooks (activate/config/mode-tracker/subagent), shell/PowerShell statusline scripts, a publish script,
__init__.py, a pi-extension, and plugin manifests for ~20 agents. The docsite import took only the markdown skill texts. Recommend keeping it that way: the executable/plugin layer is a deliberate skip (see acquisition list). - Role coverage: worker/CQR (+sweeper for debt) is the right core; runbook wiring: none (§5). Confidence: Settled.
3c. Humanizer (writing-division comparator), HELD but STALE
- Docsite:
docs/15-skills/creative/humanizer.md, blob 8318cefd68, 34,621 chars, roles [chief_of_staff]; page itself records: original author Siqi Chen (@blader), repo https://github.com/blader/humanizer version 2.5.1, MIT, Hermes additions = patterns 30–34 + a clichés list. - Upstream now: version 3.1.0 (release dated 28 Sep), ~54,533★, 97 commits, patterns rebuilt across 2.x→3.x (33→35→25→26) and a no-fabrication rule added in 2.9.0. The fleet copy is materially behind; an update must re-apply the Hermes additions on top of 3.1.0, not overwrite blindly. Upstream ships one script (scripts/validate-package.py, packaging validation only).
4. Ecosystem survey (each read live, access 2026-10-07)
- agentskills.io, the Agent Skills open standard home (originated by Anthropic; developed via agentskills/agentskills + Discord). Spec: folder + SKILL.md (required name/description; optional license/compatibility/metadata/allowed-tools; scripts/references/assets allowed; progressive disclosure). It runs a client showcase (~40 clients), NOT a skills directory, and states no skill count, a standard to conform to, not a source to import from.
- anthropics/skills, Anthropic's official repo, ~180k★, updated 2026-10-05, no releases; no single repo license, many skills Apache-2.0, document skills (docx/pdf/pptx/xlsx) source-available, per-skill LICENSE.txt governs. 19 skills; index lookup shows the fleet already holds 17 of 19 in some form. Absent: frontend-design (the gap that matters), academy-guide (Anthropic-Academy-specific) and discernment-nudge (internal-facing). All 9 probed names (frontend-design, webapp-testing, doc-coauthoring, skill-creator, web-artifacts-builder, mcp-builder, internal-comms, canvas-design, algorithmic-art) are present upstream.
- skills.sh, Vercel Labs' Agent Skills Directory, paired with the vercel-labs/skills CLI (MIT, ~33k★); leaderboard ranked by installs. Signals: find-skills #1 (3.7M installs), frontend-design 959K, skill-creator 399K, superpowers' brainstorming 383K / systematic-debugging 281K / writing-plans 266K, i.e. the ecosystem's centre of gravity matches skills the fleet mostly already holds.
- GitHub Topics, agent-skills: 29,861 repos (top: anthropics/skills 180k, ponytail 157k, addyosmani/agent-skills 102k, nexu-io/open-design 100k, ComposioHQ 76k); claude-skills: ~10,254 repos. Discovery aids, star counts measure fame, not fit; every candidate below was verified at its repo.
- ComposioHQ/awesome-claude-skills, Apache-2.0 badge (no root LICENSE file in the listing; per-skill caveat), ~76,643★, created 2025-10-17, claims 1000+ skills/plugins, ~22 bundled folders. The fleet already holds a 25-page composio category (incl. content-research-writer, lead-research-assistant), so this is a known, partially-imported source.
- obra/superpowers, Jesse Vincent (obra) + Prime Radiant, MIT, ~296k★, 683 commits, 14 releases, latest v6.4.2 (2026-09-25). 15 skills, index lookup: all 15 are already in the docsite index (brainstorming, executing-plans, writing-plans, systematic-debugging, test-driven-development, requesting/receiving-code-review, verification-before-completion, dispatching-parallel-agents, subagent-driven-development, using-git-worktrees, finishing-a-development-branch, using-superpowers, writing-skills; and skill-diagnosing-bugs as the counterpart of upstream diagnosing-superpowers). Trust note: the repo ships hooks/scripts and its brainstorming visual companion loads a Prime Radiant logo remotely (telemetry, opt-out via SUPERPOWERS_DISABLE_TELEMETRY), again, markdown-text import only.
- Others verified: VoltAgent/awesome-agent-skills (MIT, ~35k★, actively curated, claims 1000+), a good watch-list source; addyosmani/agent-skills (~102k★), the fleet already has a sweeper-tagged cron skill
skill-cron-sync-addyosmani-agent-skills-230068, i.e. a sync relationship already exists; travisvn/awesome-claude-skills (~15k★, NO license shown, README stale since Nov 2025), skip as a source; hesreallyhim/awesome-claude-code (~55k★, license Other/NOASSERTION, broader than skills), watch only.
5. Utilization findings (Q6), the wiring gap is the finding
Sampled live from jknash/agent-orchestration main: chief-of-staff.md (blob 8d6b7fd66d), worker.md (e2e24c1e05), code-quality-reviewer.md (9582617321), ui-ux.md (0796aea20b), planner.md (b034f27037), security-reviewer.md (0403637b72), lead-researcher.md (c64076858b), research-assistant.md (5da9930208). Greps run on the fetched texts (auditable):
ponytailin all 8 runbooks: 0 hits.no-ai-slop/ai-slop/humanizer: 0 hits.- The word "skill" outside the two research runbooks appears only once, in an unrelated sense (CQR's "a manifest of files and digests"). Only lead-researcher.md and research-assistant.md have an Artifacts "Skills of the role" section listing their manifest skills; chief, worker, CQR, ui_ux, planner, and security runbooks name no skills at all, they rely on bootstraps pointing at the skills library in general.
- Divergence, stated plainly: index tags say no-ai-slop→chief and ponytail→worker/CQR; manifests agree (both skills are in those roles' manifest files); runbooks say nothing. An agent working strictly from its runbook would never discover either named skill. The ui_ux case is triple: the desk's runbook names no skills, the index tags it zero skills, and no ui_ux manifest exists for the AF-46 machinery to hand it.
- Secondary: the index's ~8 anthropic/claude duplicate slots (§2) mean manifest counts overstate distinct capability by ~2%.
6. Agile project management skills (Amendment 1, added 2026-10-07)
The owner added this area mid-mission, aimed mainly at the scrum master (planner and chief of staff secondary). Every candidate below was run through one fit test: the fleet's method closes sprints when committed scope completes and never on a timebox, runs back-to-back including weekends, boards work on the in-repo llm-kanban, sizes stories in Fibonacci points under a two-stage backlog-to-story intake, and reserves acceptance to the owner. A candidate earned acquire only if it strengthens that method. Candidates built on timeboxed cadence or on Jira/Linear ceremony were marked acquire-adapted (adaptation named) or skip (conflict named).
Baseline, as found
- Scrum manifest on docsite main (
docs/15-skills/manifests/scrum_master.json, blob b0cb84e235): exactly 18 entries, skill-retro, skill-handoff, skill-claude-handoff, skill-to-tickets, skill-triage, skill-github-issues, skill-meeting-action-items, skill-sdlc-review, skill-verification-before-completion, skill-changelog-generator, skill-fleet-build-status-reporting, skill-filesystem-context, skill-using-agent-skills, skill-using-superpowers, skill-setup-matt-pocock-skills, skill-contract-registry-work-package-remediation, skill-live-delivery-reconciliation, skill-repo-and-tracker-consolidation. Character: fleet-operations skills (handoffs, triage, tickets, board/status reporting) plus one ceremony skill (retro). It is an operations set, not an agile-method set. - Discrepancy, named:
skill-planning-and-task-breakdown(docs/15-skills/engineering/planning-and-task-breakdown.md) is tagged [planner] only in the live index and is absent from the scrum manifest on main, notwithstanding the chief's disposition this morning adding it to the scrum set. State as found: disposition recorded, manifest not yet regenerated. Wire-in row 16 below carries the fix. - Vocabulary probe (redone properly; query = term searched across id+title+purpose of all 429 index entries): "sprint" 0, "groom" 0, "standup" 0, "velocity" 0, "agile" 0, "definition-of" 0, "scrum" 0, "user stor" 0. "retro" 2, skill-retro plus one incidental purpose-text hit (skill-banner-design). "estimat" 4, skill-planning-and-task-breakdown is the only method-shaped hit; the rest are incidental. "backlog" 2, both incidental (contract-registry remediation; a launcher-recovery skill). The lead's probe reproduces exactly.
- RUNBOOKS.md pointer, resolved:
~/workspace/fleet-agents/RUNBOOKS.md(scrum master table) points "Sprint / retro method" at "docsitedocs/15-skills/planning & agile method set". No such set exists as a named grouping. A full docsite git-tree scan finds zero paths containing "agile" under docs/15-skills and only two containing "planning" (engineering/planning-and-task-breakdown.md,productivity/weekly-review-planning.md); there is no planning or agile category folder among the 15-skills subdirectories. What the pointer can honestly resolve to is three scattered singles: skill-retro (ceremony), skill-planning-and-task-breakdown (task breakdown, planner-tagged), skill-weekly-review-planning (a review cadence, not sprint method). Judgement: the pointer overstates the depth behind it, it names a "method set" where the library holds, at most, one directly on-point skill (retro) and two adjacents. Recommend the pointer be corrected or the set be built by the wire-in/acquire rows below; that decision is the chief's.