All skills
lukemurraynz avatar

/azure-sre-agent

@2cc2455

Design, configure, review, and operate production-grade Azure SRE Agent capabilities: response plans, scheduled tasks, HTTP triggers, custom agents, autonomous and review workflows, approval guardrails, AMBA observability, source RCA, connectors, MCP, governance hooks, WAF reviews, AI Foundry posture, Digital Native governance, postmortem generation, and KT discipline.

Use this Skill: https://skilld.dev/gh/lukemurraynz/hve-agent-skills/azure-sre-agent

This session only. Nothing lands on disk.

bundlesobservability-ambaguide.md

≈362 tokens on demand. Your agent reads this file only when SKILL.md points to it.

AMBA Observability Bundle

Use Azure Monitor Baseline Alerts as the reference baseline for new Azure SRE Agent monitoring designs. Keep AMBA as a source-driven baseline rather than copying alert tables into prompts.

Design Rules

  1. Start with current AMBA service guidance for the Azure resource type.
  2. Classify alerts before creating response plans:
    • investigate automatically
    • notify only
    • digest/report only
    • candidate for later autonomy
    • suppress or tune after evidence
  3. Map only alerts with a useful investigation path to SRE Agent response plans.
  4. Tune thresholds to workload SLOs, business criticality, and known customer baselines.
  5. For policy-deployed AMBA alerts, check assignment scope, effect, identity, compliance state, exclusions, overrides, and action-group routing.
  6. Default AMBA-triggered response plans to Review mode. Autonomous mode requires measured alert quality and safe, repeatable mitigation.

Useful Pairings

  • Service Health and Resource Health: stakeholder impact assessment and subscription/service mapping.
  • Activity Log delete/update alerts: change validation, blast-radius review, rollback recommendation, and KT DA/PPA for production impact.
  • Metric alerts: resource-specific triage, correlation, scaling/remediation proposal, and threshold tuning feedback.
  • Log-search alerts: fleet investigation, KQL summarisation, recurring pattern detection, and knowledge capture.

Source: SKILL.md on GitHub

No alerts8d3 checks · Risk SAFE
  • Gen Agent Trust Hub8d

    The Azure SRE Agent skill provides a production-grade framework for managing Azure infrastructure using AI agents. It incorporates extensive safety documentation, approval-based hooks, and least-privilege role templates. The 'low' verdict is assigned due to the inherent risk of indirect prompt injection when the agent processes external incident data and source code, a necessary function for its SRE capabilities.

  • Socket8d

    No alerts

  • Snyk8d

    Risk: LOW · No issues

Signed by skilld at 2cc2455. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub last month.

Steadyupdated last month
compatibility
Azure SRE Agent; GitHub Copilot agent skills; new projects only
Other metadata
metadata
{
  "last_verified": "2026-08-25",
  "version": "2.23.3",
  "risk": "critical",
  "last_updated": "2026-08-25"
}

README badge

README badge for lukemurraynz/hve-agent-skills/azure-sre-agent