All skills
mblode avatar

/agents-md

@59baecd
by Matthew Blodemblode/agent-skills136 stars
12

Audits and edits agent instruction files, verifies repository commands, and migrates repositories to AGENTS.md as the single shared source. Use when asked to "improve my AGENTS.md", "migrate CLAUDE.md to AGENTS.md", or make instructions work across agents. For SKILL.md use agent-skills-creator.

Use this Skill: https://skilld.dev/gh/mblode/agent-skills/agents-md

This session only. Nothing lands on disk.

referencesquality-criteria.md

≈1.3k tokens on demand. Your agent reads this file only when SKILL.md points to it.

Full Quality Criteria (49 Checks)

Score each AGENTS.md root file against this checklist. Standard: the file helps an agent execute correctly with minimal context.

Contents

  • Scoring
  • A. Commands and execution readiness (12 checks)
  • B. Gotchas and repeated mistakes (10 checks)
  • C. Conventions and decision boundaries (11 checks)
  • D. Signal-to-noise and bloat control (9 checks)
  • E. Currency and validation (7 checks)
  • Grade mapping
  • Automatic fails

Scoring

  • Yes = 1, No = 0, N/A = excluded from the denominator
  • Grade uses earned / applicable
  • Target: >= 91% of applicable points (grade A)

A. Commands and execution readiness (12)

  1. Working dev command (or equivalent local run)
  2. Working test command
  3. Working build command
  4. Working lint and/or typecheck command
  5. Deploy/release command when applicable
  6. Migration/seed/db command when applicable
  7. Commands are copy-paste ready (no placeholders)
  8. Commands match the package manager and scripts
  9. Required env bootstrap steps (incl. secondary runtimes like Python venvs)
  10. Where to run commands (root/workspace)
  11. A command for targeted test/debug iteration, quiet flags included so a full passing suite does not linger in later turns
  12. No duplicate or conflicting command variants

B. Gotchas and repeated mistakes (10)

  1. At least one high-frequency failure mode
  2. Gotchas are project-specific, not generic
  3. Gotchas include corrective action (what to do instead)
  4. Gotchas include trigger context (when the rule applies)
  5. Captures at least one issue discovered from PR/review feedback
  6. Ordering/dependency gotchas where order matters
  7. Data/env gotchas where setup mistakes cause failures
  8. Avoids vague advice like "be careful" or "follow patterns"
  9. Separates universal rules from edge-case rules
  10. Removes gotchas that no longer happen

C. Conventions and decision boundaries (11)

  1. States conventions that materially change implementation choices
  2. States naming/path conventions when CI/tooling depends on them
  3. States test strategy conventions (unit/e2e boundaries) when relevant
  4. Every rule the whole repo must obey is inline in root; nothing must-obey lives only behind an @import, a .claude/rules/ file, or a .cursor/rules/ file that a single tool resolves
  5. Marks scope boundaries: monorepo root vs workspace files
  6. Avoids restating what the agent or harness already does (tool-use conventions, read before edit, todo tracking, run tests after a change)
  7. Rules name the condition that triggers them, so precision lands on when a rule applies rather than on forbidding a whole class of action
  8. Emphasis markers (IMPORTANT, NEVER, YOU MUST) used sparingly on critical rules agents skip
  9. Guidance states the outcome wanted; blanket prohibitions appear only where the harmful-precision test clears them
  10. No rule contradicts a parent instruction file, an installed skill, or another section of the same file; precedence is stated where overlap is deliberate
  11. Conventions with an exemplar in the repo name that file path instead of describing the pattern in prose

D. Signal-to-noise and bloat control (9)

  1. Root file concise for repo complexity (60-150 lines for active app repos; 200 is Claude Code's stated ceiling, and Codex stops reading at 32 KiB across all instruction files)
  2. No full framework documentation pasted inline
  3. No copy-pasted full templates
  4. No exhaustive file tree or "every file" inventory
  5. No long architecture deep dives in root file
  6. Non-universal guidance lives where it loads on demand (nested file, path-scoped rule, skill, or plain link), not in root and not behind an @import that loads at launch anyway
  7. No duplicate guidance across sections
  8. No content auto-memory owns (user preferences, personal feedback, evolving project status)
  9. Each section passes the litmus test: removing it would cause mistakes

E. Currency and validation (7)

  1. Referenced file paths exist
  2. Referenced tools/dependencies are still in use
  3. Commands have been run (or limitations documented when run isn't possible)
  4. Removed references to deleted folders/APIs
  5. Version-sensitive guidance is date/version scoped where needed
  6. Clear maintenance loop (how to keep the file current)
  7. Personal overrides stay private; any CLAUDE.local.md fallback blocker is identified and its loading mode verified

Grade mapping

Use earned / applicable percentage:

  • A: >= 91%
  • B: 76% to < 91%
  • C: 59% to < 76%
  • D: 39% to < 59%
  • F: < 39%

Example: 36/40 = 90% -> Grade B.

Automatic fails

Mark grade F regardless of score if any hold:

  • Commands are mostly broken/stale
  • Instructions are primarily generic advice, or restatements of default agent behavior
  • File is dominated by copied docs/templates rather than executable guidance
  • The intended tool does not load the shared instructions: check Claude Code version, built-in mod, Project instructions mode, and leftover project Claude files; absence of a CLAUDE.md wrapper is not a failure

Source: SKILL.md on GitHub

2 warnings4d5 checks · Risk SAFE
  • Gen Agent Trust Hub4d

    The skill is a utility for managing and auditing AI agent instruction files. It is generally safe but possesses a surface for indirect prompt injection as it processes instruction files that could contain malicious directives. It also facilitates shell command execution for repository auditing and validation.

  • Socket4d

    No alerts

  • Snyk4d

    Risk: LOW · No issues

  • Runlayer7mo

    6/6 files flagged

  • ZeroLeaks5mo

    1 finding · Score: 69/100

Signed by skilld at 59baecd. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub 27 minutes ago.

Activeupdated last week

README badge

README badge for mblode/agent-skills/agents-md