ktg-plugin-marketplace

Author	SHA1	Message	Date
Kjell Tore Guttormsen	fc8808d6e4	docs(humanizer): v5.1.0 release notes across plugin + marketplace docs - Plugin README: add "What's New in v5.1.0" section with humanizer overview, before/after example, plain-language vocabulary table, --raw flag docs. Bump version badge 5.0.0 → 5.1.0. Add Version History row. - Plugin CLAUDE.md: add humanizer.mjs + humanizer-data.mjs to Scanner Lib table. Add "Plain-Language Output (v5.1.0)" section documenting output modes, vocabularies, and Wave 5 lessons. Bump test count 635 → 792 across 52 test files. - Marketplace root README: bump config-audit entry 5.0.0 → 5.1.0, update one-line description to mention plain-language UX, add bullet for the v5.1.0 humanizer, bump test count 635+ → 792+. Test-normalizer hardening (consequence of growing CLAUDE.md): walkClaudeMdCascade walks upward from the marketplace-medium fixture into this plugin's own CLAUDE.md, so any docs edit ripples into `scanners[].activeConfig.claudeMdEstimatedTokens`. The v5.0.0 byte-stability contract is about scanner internals being unchanged, not ancestor input content being frozen. Normalizers in json-backcompat, raw-backcompat, posture-humanizer, scan-orchestrator-humanizer, and snapshot-default-output now strip claudeMdEstimatedTokens to <ANCESTOR_DERIVED>. The default-output snapshot for scan-orchestrator was re-seeded via UPDATE_SNAPSHOT=1 (intent: Wave 6 docs additions; humanizer prose unchanged). Verify: - grep -E "5\.1\.0\|v5\.1\.0" README.md CLAUDE.md ../../README.md \| wc -l = 12 - node --test 'tests//.test.mjs' = 792/792 pass - self-audit configGrade A (97), pluginGrade A (100), readmeCheck.passed true	2026-05-01 20:35:24 +02:00
Kjell Tore Guttormsen	c5c937e94e	feat(humanizer): forbidden-words lint runner + test wrapper (SC-3) [skip-docs] Step 8 of v5.1.0 humanizer Wave 4. Adds tests/lint-default-output.mjs runner and tests/scanners/lint-default-output.test.mjs wrapper that exercise SC-3 against the 6 prose CLIs (scan-orchestrator, posture, token-hotspots-cli, plugin-health-scanner, drift-cli, fix-cli) running in default (humanized) mode against tests/fixtures/marketplace-medium. Lint scope is stderr only — JSON envelope keys ("scanner", "severity") are structural, not prose. Humanized prose fields embedded inside JSON are already covered by tests/lib/humanizer-data.test.mjs tier1/tier3 checks. Code references inside backticks pass the lint (stripBacktickSpans) so technical identifiers can appear when wrapped. Default-mode prose fixes to land lint at zero violations: - scan-orchestrator: top banner switches to "Config-Audit v2.2.0" and per-scanner progress wraps "[XXX] Label" in backticks. --raw and --json paths preserve the v5.0.0 verbatim banner via new opts.humanizedProgress flag on runAllScanners. - plugin-health-scanner: top banner switches to "Plugin Health v2.1.0" in default mode; --raw/--json keep "Plugin Health Scanner v2.1.0". - scoring.mjs generateHealthScorecard humanized branch: area names (CLAUDE.md, Hooks, MCP, Settings, Rules, Imports, Conflicts, Token Efficiency, Plugin Hygiene) are wrapped in backticks; dot-padding compensates so column alignment matches v5.0.0 layout. - posture / drift-cli / fix-cli: thread humanizedProgress flag through their runAllScanners calls so default mode emits humanized progress and --raw/--json preserve the v5.0.0 stderr snapshot. Test infrastructure only — user-facing docs land in Wave 5/6 once commands and agents consume the humanized payload. Tests: 735 to 736 (+1 SC-3 wrapper). Full suite passes. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 18:11:15 +02:00
Kjell Tore Guttormsen	5eecb968d8	feat(humanizer): wire humanizer into 6 remaining CLIs with --raw Adds --raw flag to all 6 remaining CLIs and wires humanization into the default rendering path. --json and --raw both bypass humanization for v5.0.0 byte-equal output; default mode humanizes findings/diff/prose. token-hotspots-cli: humanizes payload.findings before stdout JSON write. plugin-health-scanner: humanizes finding titles in stderr brief summary; --json/--raw write byte-identical v5.0.0-shape result to stdout. drift-cli: humanizes diff.{newFindings,resolvedFindings,unchangedFindings, movedFindings} before formatDiffReport; --raw applies to save and list modes too. Baselines remain raw v5.0.0 on disk. fix-cli: humanizes manual-finding titles in stderr fix-plan prose; both --json and --raw produce identical machine-readable JSON to stdout. manifest, whats-active: --raw is a no-op (no findings, inventory only) but parsed for CLI surface consistency. Decision on missing --output-file flag for drift-cli/fix-cli/plugin-health: deferred. SC-6/SC-7 tests in Wave 4 will use stdout-redirect (the simpler Alt B path) since these CLIs already write JSON to stdout in machine modes. Test cli-humanizer.test.mjs covers all 6 CLIs. Three CLIs that read environment state (plugin-health, manifest, whats-active) verify mode-equivalence (--json == --raw) instead of frozen-snapshot byte-equal, because their output reflects current marketplace state which drifts as plugins are added since the Wave 0 capture. Wave 3 / Step 7 of v5.1.0 humanizer. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 17:47:09 +02:00
Kjell Tore Guttormsen	70ff900578	feat(humanizer): wire humanizer into posture and scoring scorecard generateHealthScorecard signature: 2-arg → 3-arg (areaScores, opportunityCount, options = {}). options.humanized=true renders friendlier title, grade-context line per overall grade, and rephrased opportunity line. options.humanized=false (or 2-arg call) preserves v5.0.0 verbatim output for backwards-compat. topActions also gets an optional options.humanized that swaps recommendations through humanizeFinding lookup. posture.mjs main(): --json → write JSON to stdout, suppress stderr scorecard --raw → write JSON to stdout (byte-identical to --json), write v5.0.0 verbatim scorecard to stderr default → humanized scorecard to stderr, no stdout posture.test.mjs scorecard-prose assertions re-anchored to --raw mode (the explicit v5.0.0 path) — Wave 0 audit only covered finding-title strings; scorecard prose surfaces here for the first time. Wave 3 / Step 6 of v5.1.0 humanizer. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 17:38:03 +02:00
Kjell Tore Guttormsen	5ff6594976	feat(humanizer): wire humanizer into scan-orchestrator main with --raw bypass Adds --json and --raw flags to scan-orchestrator CLI main(). Default mode runs humanizeEnvelope(env) before serialization; --json and --raw bypass the humanizer for v5.0.0 byte-equal output (SC-6 / SC-7 paths). Save-baseline path always writes the raw v5.0.0-shape envelope so future humanizer-data updates do not trigger false-positive drift findings. runAllScanners() unchanged — it remains the v5.0.0-shape source of truth for in-process callers (posture, scoring, drift, etc.). Wave 3 / Step 5 of v5.1.0 humanizer. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 17:31:37 +02:00
Kjell Tore Guttormsen	dff278f02a	test(humanizer): replace title-string assertions with ID-based checks Wave 2 / Step 4 of v5.1.0 plain-language UX humanizer rollout. Re-anchors 34 title-string assertions across 4 test files so they survive Wave 3's title/description/recommendation rewriting at the CLI layer. Anchoring strategy per scanner: - GAP findings: scanner + category + recommendation substring (humanizer preserves stable identifiers like CLAUDE.md, .mcp.json, hook in rec). Hardcoded CA-GAP-NNN IDs for positive checks. - HKV findings: scanner + evidence regex (evidence preserved verbatim). - SET findings: scanner + evidence regex (evidence preserved verbatim). - PLH findings: scanner + hardcoded CA-PLH-NNN IDs (no evidence on most PLH findings, so ID is the only stable anchor for specific cases; negative checks use scanner + title-substring spanning raw + humanized). Per docs/v5.1.0-test-audit.md classification: only (b) WILL BREAK assertions modified. (a) shape-only assertions (error-message formatting, pure existence checks) untouched. tests/lib/output.test.mjs and tests/lib/diff-engine.test.mjs and tests/scanners/fix-engine.test.mjs unchanged (synthetic test inputs, not scanner output). Test count unchanged: 689/689 pass. IDs harvested via deterministic runtime dump per fixture (resetCounter + scan).	2026-05-01 17:22:55 +02:00
Kjell Tore Guttormsen	b7414303de	feat(config-audit): --accurate-tokens API calibration (v5 N5) [skip-docs]	2026-05-01 09:15:02 +02:00
Kjell Tore Guttormsen	df6e012903	docs(config-audit): cache-telemetry recipe + --with-telemetry-recipe flag (v5 M7)	2026-05-01 09:12:17 +02:00
Kjell Tore Guttormsen	cd25c1e934	feat(config-audit): cross-plugin collision scanner COL (v5 N6) [skip-docs] New COL scanner detects skill-name collisions across plugins and between user-level skills (~/.claude/skills/) and plugin-bundled skills. Skill identity is the directory basename — matches how enumerateSkills resolves names. Detection rules (per docs/v5-namespace-research.md, confidence: medium): - Plugin-vs-plugin same skill name → severity low (CA-COL-001) - User-vs-plugin same skill name → severity medium (CA-COL-001) - Plugin-vs-built-in collisions: out of scope for v5.0.0 (insufficient verification — recorded for v5.0.1 follow-up). Findings carry details.namespaces array with {source, name, path} for every conflicting source — supports per-collision reporting downstream. output.mjs: finding() helper now passes through optional `details` field (scanner-specific structured payload). scoring.mjs: COL → "Plugin Hygiene" (new area, 10 total). Posture test updated from 9 → 10 area scores. .gitignore: docs/v5-namespace-research.md is local-only (Step 22a research output, gitignored per plan). Fixture collision-plugins/fake-home/ has user skill `review` colliding with plugin-a + plugin-b's `review` (medium severity), plus plugin-c's unique `summarize` (no collision). [skip-docs] reason: v5 plan fences off README/CLAUDE.md badge updates to Session 5; Forgejo pre-commit-docs-gate hook requires this tag. Tests: 617 → 625 (+8). Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 07:46:15 +02:00
Kjell Tore Guttormsen	cc349d6fe1	feat(config-audit): disabled-in-schema scanner DIS (v5 N4) [skip-docs] New DIS scanner detects tools that appear in BOTH permissions.deny and permissions.allow within the same settings.json file. The deny list wins, so allow entries are dead config but still load on every turn and confuse intent. Tool identity = bare name (everything before "("). `Bash(npm:*)` and `Bash` are treated as the same tool, so a deny on `Bash` flags any `Bash(...)` allow entry. Severity: low. Wired into scan-orchestrator + scoring (area: Settings). Fixture denied-tools-in-schema has Bash in both arrays; healthy-project serves as the negative case. [skip-docs] reason: v5 plan fences off README/CLAUDE.md badge updates to Session 5; Forgejo pre-commit-docs-gate hook requires this tag. Tests: 611 → 617 (+6). Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 07:39:58 +02:00
Kjell Tore Guttormsen	65087e624f	feat(config-audit): cache-prefix stability scanner CPS (v5 N3) [skip-docs] New CPS scanner walks CLAUDE.md cascade and flags volatile content between lines 31 and 150 — the cache-prefix window beyond TOK Pattern A's top-30 territory. Volatile content anywhere in the cached prefix forces a fresh cache write from that line down on every turn. Volatile-pattern set extends TOK Pattern A with: - shell-exec lines (! prefix) — common in CLAUDE.md to inject git/date - ${VAR} substitutions — vary per-shell, defeat cache reuse Severity: medium per finding. Skips lines 1-30 to avoid duplicating Pattern A's range; CPS' value is in the 31-150 zone. Wired into scan-orchestrator + scoring SCANNER_AREA_MAP. CPS shares the "Token Efficiency" area with TOK; scoreByArea now deduplicates by area name and combines counts across scanners contributing to the same area, so the 9-area scorecard contract holds. Fixtures volatile-mid-section/{volatile-line-60, volatile-line-200} verify both positive (line 60) and out-of-window (line 200) cases. [skip-docs] reason: v5 plan fences off README/CLAUDE.md badge updates to Session 5; Forgejo pre-commit-docs-gate hook requires this tag. Tests: 604 → 611 (+7). Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 07:37:54 +02:00
Kjell Tore Guttormsen	0420b8cc4a	feat(config-audit): /config-audit manifest command (v5 N2) [skip-docs] New scanners/manifest.mjs CLI + commands/manifest.md slash command. Reads activeConfig and produces a flat, ranked list of every token source (CLAUDE.md cascade entries, plugins, skills, MCP servers, hooks) sorted DESC by estimated_tokens. CLAUDE.md per-file tokens are derived by distributing claudeMd.estimatedTokens across the cascade proportional to bytes. Tests cover both real-config (plugin root) and fixture (rich-repo with patched HOME containing 2 plugins + 3 skills + .mcp.json) paths, plus error handling (nonexistent path → exit 3, --output-file). Builds on readActiveConfig from M1 (v5 alpha.2). [skip-docs] reason: v5 plan fences off README/CLAUDE.md badge updates to Session 5; Forgejo pre-commit-docs-gate hook requires this tag on feat commits without doc changes. Tests: 593 → 604 (+11). Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 07:32:54 +02:00
Kjell Tore Guttormsen	b2407a09b3	feat(config-audit): CA-TOK-005 MCP tool-schema budget (v5 N1) [skip-docs] Adds detectMcpToolBudget detection block in TOK scanner. Tiered severity per project-local .mcp.json server based on toolCount: - < 20: no finding - 20-49: low - 50-99: medium - 100+: high - null (manifest unparseable): low + "tool count unknown" message Scoped to source==='.mcp.json' to keep findings actionable for the audited path; plugin/user-level MCP servers are surfaced by the manifest scanner (Step 19 / N2). 5 fixtures (mcp-budget/{14,25,60,120,unknown}-tools) use inline `tools` arrays in .mcp.json — no node_modules needed for these tests. Tests assert title+severity (not exact ID) since TOK IDs are sequential per scan, not semantic per pattern. [skip-docs] reason: v5 plan fences off README/CLAUDE.md badge updates to Session 5; Forgejo pre-commit-docs-gate hook requires this tag on feat commits without doc changes. Tests: 586 → 593 (+7). Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 07:29:57 +02:00
Kjell Tore Guttormsen	3c79f95e9a	feat(config-audit): self-audit --check-readme flag (v5 F6) [skip-docs] Filesystem counts are the source of truth; README badges parsed via line-anchored substring (badge/<kind>-<N>-...). Emits readmeCheck object with counts/badges/mismatches. CLI: node scanners/self-audit.mjs --check-readme [--json] API: runSelfAudit({ checkReadme: true }) → result.readmeCheck Helper: checkReadmeBadges(pluginDir) for per-fixture testing New fixture: readme-desynced/ (commands/foo + bar, README claims 1). Note: alpha phase does NOT require result.readmeCheck.passed === true. Self-test of real plugin currently fails (scanners 10 vs 9, tests 31 vs 543); will be reconciled in Session 5 Step 28 (README sync). 582 → 586 tests, all green.	2026-05-01 07:09:26 +02:00
Kjell Tore Guttormsen	910567d661	feat(config-audit): HKV flags verbose hook output (v5 M5) [skip-docs] Static heuristic — counts console.log / process.stdout.write lines per referenced hook script. > 50 → low CA-HKV-NNN finding. New fixtures: - hooks-verbose/ (61 verbose lines → triggers) - hooks-quiet/ (5 lines → no finding) 580 → 582 tests, all green.	2026-05-01 07:05:45 +02:00
Kjell Tore Guttormsen	9a44df22ac	feat(config-audit): TOK flags skill description > 500 chars (v5 M2) [skip-docs] - New Pattern F in TOK: low-severity finding when SKILL.md description > 500 chars - Scoped to discovery.files (project-local) — activeConfig.skills walk would pull in user/plugin skills out of project scope - New fixtures: skill-bloated (594-char desc) + skill-tight (46-char baseline) 574 → 576 tests, all green.	2026-05-01 06:58:42 +02:00
Kjell Tore Guttormsen	25ca6139b4	feat(config-audit): TOK flags CLAUDE.md cascade > 10k tokens (v5 M4) [skip-docs] - New Pattern E in TOK: emits medium finding when activeConfig.claudeMd.estimatedTokens > 10_000 - Uses cascade tokens, file count, and calibration note as evidence - New fixtures: large-cascade (37k bytes / 14475 cascade tokens) + small-cascade (5k baseline) 572 → 574 tests, all green.	2026-05-01 06:53:12 +02:00
Kjell Tore Guttormsen	9330124f5c	feat(config-audit): flag additionalDirectories > 2 (v5 M6) [skip-docs] - Add 'additionalDirectories' to KNOWN_KEYS - Emit low severity finding when length > 2 - New fixtures: additional-dirs-many (3 entries) + additional-dirs-ok (2) 569 → 572 tests, all green.	2026-05-01 06:50:24 +02:00
Kjell Tore Guttormsen	58d6b5b9ea	feat(config-audit): recalibrate TOK severities for tokens/turn (v5 F7) [skip-docs] - Pattern A (cache-breaking volatile top): medium → high - Pattern B (redundant permissions): low → medium - Pattern C (deep @import chain): medium → low - Add calibration_note evidence on every TOK finding - Table-driven severity tests (identify by title, IDs are sequential) 563 → 569 tests, all green. Doc sweep deferred to Session 5 (Step 28).	2026-05-01 06:47:32 +02:00
Kjell Tore Guttormsen	2810ee6f62	feat(config-audit): remove TOK Pattern D detectSonnetEra (v5 F5) Pattern D was the v4 sonnet-era signature: 'config is structurally clean but uses no Opus-4.7-specific features'. Two problems: - It triggered on any minimal config that happened to lack skills/MCP - The advice was generic and not actionable The hotspots ranking and per-pattern findings (A/B/C) cover the same ground with concrete, file-anchored signal. Dropping the noise. BREAKING (intentional): scanners no longer emit the sonnet-era info finding. Suppression entries and downstream tooling that reference the v4 finding ID should be updated. Doc sweep follows in Step 8b. Tests: sonnet-era fixture now asserts zero findings.	2026-05-01 06:31:43 +02:00
Kjell Tore Guttormsen	0d8a9af3d6	fix(config-audit): remove TOK dead take + hotspot padding (v5 F4) The buildHotspots padding loop and unused 'take' variable were dead code from the v3 hotspots-min contract. Replaced with a clean ranked.slice(0, HOTSPOTS_MAX). Tiny fixtures may now return fewer than 3 hotspots, which is the honest answer; the contract now only asserts <= 10. Tests: +2 cases — every hotspot.source is unique (no padding); length never exceeds HOTSPOTS_MAX.	2026-05-01 06:29:33 +02:00
Kjell Tore Guttormsen	34669d596c	feat(config-audit): TOK consumes readActiveConfig (v5 F1) Removes the v4 'void readActiveConfig' placeholder and wires the active-config snapshot into the TOK scanner. Per-turn behavior changes: - Each enabled MCP server becomes its own hotspot entry (richer than the parent .mcp.json file alone) - total_estimated_tokens now includes MCP server cost - result.activeConfig exposes a small summary (claudeMdEstimatedTokens, mcpServerCount, pluginCount, skillCount) Failures of readActiveConfig are non-fatal — the scanner falls back to the discovery-only path used in v4. Tests: +3 cases on the new tok-active-config fixture (.mcp.json with 2 servers, CLAUDE.md, plugin skeleton).	2026-05-01 06:27:34 +02:00
Kjell Tore Guttormsen	cbc889f603	feat(config-audit): add token-hotspots CLI (node scanners/token-hotspots-cli.mjs)	2026-04-19 22:46:25 +02:00
Kjell Tore Guttormsen	295a6289b4	test(config-audit): extend grade-stability test to assert Token Efficiency A/B on baseline	2026-04-19 22:45:34 +02:00
Kjell Tore Guttormsen	4b385bf456	feat(config-audit): wire TOK into posture scorecard as 8th quality area (Token Efficiency)	2026-04-19 22:45:12 +02:00
Kjell Tore Guttormsen	712058c387	test(config-audit): add token-hotspots scanner test suite (red)	2026-04-19 22:40:44 +02:00
Kjell Tore Guttormsen	350cebc39c	test(config-audit): add baseline-all-a fixture + grade-stability regression test	2026-04-19 22:32:40 +02:00
Kjell Tore Guttormsen	f93d6abdae	feat: initial open marketplace with llm-security, config-audit, ultraplan-local	2026-04-06 18:47:49 +02:00

28 commits