All skills
simota avatar

/judge

@e307415
by shingo imotasimota/agent-skills85 stars
15

Reviewing code via multi-engine orchestration (Claude + Codex) on three axes — secure, correct, and lean — shipping only findings worth fixing. Use for PR review or pre-commit. Complements Zen.

Use this Skill: https://skilld.dev/gh/simota/agent-skills/judge

This session only. Nothing lands on disk.

referenceautorun-schema.md

≈409 tokens on demand. Your agent reads this file only when SKILL.md points to it.

Judge — AUTORUN _STEP_COMPLETE Schema

When Judge receives _AGENT_CONTEXT, parse task_type, description, review_mode, base_branch, and Constraints, choose the review mode, run the default tri-engine workflow (or single-engine fallback; lean runs it with a lean focus), and return _STEP_COMPLETE. pair mode is INTERACTIVE and cannot run unattended — under AUTORUN, perform the review/seed half and return ranked findings with Next: USER (pair-ready), never applying fixes without confirmation.

_STEP_COMPLETE

_STEP_COMPLETE:
  Agent: Judge
  Status: SUCCESS | PARTIAL | BLOCKED | FAILED
  Output:
    deliverable: [report path or inline]
    artifact_type: "[PR | Pre-Commit | Commit | Consistency | Test Quality | Lean | Pair]"
    parameters:
      review_mode: "[Tri-Engine | Single-Engine (codex|agy|claude) | Pair | GitHub-Async]"
      engines_run: "[actual usable engines]"
      engines_failed: "[list or none]"
      files_reviewed: "[count]"
      findings_shipped: "[CRITICAL: N, HIGH: N, MEDIUM: N, LOW: N, INFO: N]"
      lean_findings: "[count or N/A — waste patterns L1–L6]"
      concurrence: "[votes/usable_engines: count, with grounded singleton counts]"
      rejected: "[count + top categories]"
      verdict: "[APPROVE | REQUEST CHANGES | BLOCK]"
      intent_alignment: "[PASS | FAIL | NOT_CHECKED]"
      consistency_issues: "[count or none]"
      test_quality_score: "[score or N/A]"
      pair_outcomes: "[Pair only — RESOLVED/REJECTED/DEFERRED/REGRESSED: N | N/A]"
  Next: Builder | Sentinel | Zen | Radar | USER | DONE
  Reason: [Why this next step]

Source: SKILL.md on GitHub

No alerts13d5 checks · Risk SAFE
  • Gen Agent Trust Hub13d

    The 'judge' skill is an advanced code review tool designed to analyze software changes using multiple AI engines including Claude, Codex, and Google Gemini. It features sophisticated workflows for PR reviews, security audits, and performance checks. While it processes external data like code diffs and PR descriptions, it includes a multi-layered verification process to ensure findings are accurate. The skill also includes technical instructions for managing CLI tool requirements, such as handling terminal interactions using Python. All external resources and download instructions point to trusted organizations like Google and Anthropic.

  • Socket13d

    No alerts

  • Snyk13d

    Risk: LOW · No issues

  • Runlayer6mo

    1/11 files flagged

  • ZeroLeaks5mo

    Score: 93/100 · 2 sections analyzed

Signed by skilld at e307415. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub 2 days ago.

Activeupdated 2 weeks ago

README badge

README badge for simota/agent-skills/judge