All skills
nvidia avatar

/physical-ai-infrastructure-setup-and-resilient-scaling

@6c6fc09
by NVIDIA Corporationnvidia/skills3.5k stars
424

Use when the user wants to set up, scale, validate, or harden NVIDIA physical AI infrastructure for synthetic data generation workflows across local MicroK8s or Azure AKS, including Kubernetes clusters, inference endpoint deployment, OSMO deployment, workload submission readiness, and infrastructure failure recovery. Trigger keywords: physical ai infrastructure, resilient scaling, SDG infrastructure, microk8s, azure aks, NVCF deployment, NIM Operator, OSMO deploy, workflow scaling. Don't trigger for: OSMO log summarization or workload-only operations unless infrastructure setup, scaling, validation, or recovery is requested.

Use this Skill: https://skilld.dev/gh/nvidia/skills/physical-ai-infrastructure-setup-and-resilient-scaling

This session only. Nothing lands on disk.

BENCHMARK.md

≈1.6k tokens on demand. Your agent reads this file only when SKILL.md points to it.

Evaluation Report

Evaluation of the physical-ai-infrastructure-setup-and-resilient-scaling skill before publication through NVSkills-Eval.

This benchmark summarizes 3-Tier Evaluation from NVSkills-Eval results for the skill. The goal is to document whether the skill is safe, discoverable, effective, and useful for agents before it is published for broader workflow use.

Evaluation Summary

  • Skill: physical-ai-infrastructure-setup-and-resilient-scaling
  • Evaluation date: 2026-05-29
  • NVSkills-Eval profile: external
  • Overall verdict: FAIL
  • Tier 3 live agent evaluation: not available in this report

Agents Used

  • Tier 3 agent details were not available in this report.

Metrics Used

Reported benchmark dimensions:

  • Security: checks whether skill-assisted execution avoids unsafe behavior such as secret leakage, destructive commands, or unauthorized access.
  • Correctness: checks whether the agent follows the expected workflow and produces the correct final output.
  • Discoverability: checks whether the agent loads the skill when relevant and avoids using it when irrelevant.
  • Effectiveness: checks whether the agent performs measurably better with the skill than without it.
  • Efficiency: checks whether the agent uses fewer tokens and avoids redundant work.

Underlying evaluation signals used in this run:

  • No Tier 3 evaluation signal details were available in this report.

Test Tasks

Tier 3 evaluation task details were not available in this report.

Results

Tier 3 dimension rollup was not available in this report.

Tier 1: Static Validation Summary

Tier 1 validation passed with observations. NVSkills-Eval ran 9 checks and found 13 total findings.

Top findings:

  • MEDIUM SCHEMA/body_recommended_section: Missing recommended section: '## Instructions' (skills/physical-ai-infrastructure-setup-and-resilient-scaling/SKILL.md)
  • MEDIUM SCHEMA/body_recommended_section: Missing recommended section: '## Examples' (skills/physical-ai-infrastructure-setup-and-resilient-scaling/SKILL.md)
  • LOW QUALITY/quality_correctness: No examples provided (skills/physical-ai-infrastructure-setup-and-resilient-scaling/SKILL.md)
  • LOW QUALITY/quality_discoverability: Description very long (632 chars, recommend 50-150) (skills/physical-ai-infrastructure-setup-and-resilient-scaling/SKILL.md)
  • LOW QUALITY/quality_discoverability: Broad description without negative triggers may cause over-triggering (skills/physical-ai-infrastructure-setup-and-resilient-scaling/SKILL.md)

Tier 2: Deduplication Summary

Tier 2 validation reported findings. NVSkills-Eval ran 2 checks and found 16 total findings.

Top findings:

  • HIGH DUPLICATE/duplicate: Duplicate content found across components/osmo-azure/reference.md and components/osmo-k8s/reference.md: "# Re-run" in components/osmo-azure/reference.md (lines 98-102) vs "# Re-run" in components/osmo-k8s/reference.md (lines 75-79) (components/osmo-azure/reference.md:98)
  • HIGH DUPLICATE/duplicate: Duplicate content found across components/osmo-azure/reference.md and components/osmo-k8s/reference.md: "# Verify" in components/osmo-azure/reference.md (lines 86-89) vs "# Verify" in components/osmo-k8s/reference.md (lines 71-74) (components/osmo-azure/reference.md:86)
  • HIGH DUPLICATE/duplicate: Duplicate content found across components/cluster-azure/scripts/preflight.sh and components/cluster-microk8s/scripts/preflight.sh and components/inference-azure/scripts/preflight.sh and components/inference-nim-operator/scripts/preflight.sh and components/inference-nvcf/scripts/preflight.sh and components/osmo-azure/scripts/preflight.sh and components/osmo-cli/scripts/preflight.sh and components/osmo-k8s/scripts/preflight.sh: "check_min_version()" in components/cluster-azure/scripts/preflight.sh (lines 118-129) vs "check_min_version()" in components/cluster-microk8s/scripts/preflight.sh (lines 62-73) vs "check_min_version()" in components/inference-azure/scripts/preflight.sh (lines 110-121) vs "check_min_version()" in components/inference-nim-operator/scripts/preflight.sh (lines 47-58) vs "check_min_version()" in components/inference-nvcf/scripts/preflight.sh (lines 47-58) vs "check_min_version()" in components/osmo-azure/scripts/preflight.sh (lines 110-121) vs "check_min_version()" in components/osmo-cli/scripts/preflight.sh (lines 42-53) vs "check_min_version()" in components/osmo-k8s/scripts/preflight.sh (lines 48-59) (components/cluster-azure/scripts/preflight.sh:118)
  • HIGH DUPLICATE/duplicate: Duplicate content found across components/cluster-azure/scripts/preflight.sh and components/cluster-microk8s/scripts/preflight.sh and components/inference-azure/scripts/preflight.sh and components/inference-nim-operator/scripts/preflight.sh and components/inference-nvcf/scripts/preflight.sh and components/osmo-azure/scripts/preflight.sh and components/osmo-k8s/scripts/preflight.sh: "require_cmds()" in components/cluster-azure/scripts/preflight.sh (lines 30-39) vs "require_cmds()" in components/cluster-microk8s/scripts/preflight.sh (lines 17-26) vs "require_cmds()" in components/inference-azure/scripts/preflight.sh (lines 22-31) vs "require_cmds()" in components/inference-nim-operator/scripts/preflight.sh (lines 19-28) vs "require_cmds()" in components/inference-nvcf/scripts/preflight.sh (lines 19-28) vs "require_cmds()" in components/osmo-azure/scripts/preflight.sh (lines 22-31) vs "require_cmds()" in components/osmo-k8s/scripts/preflight.sh (lines 20-29) (components/cluster-azure/scripts/preflight.sh:30)
  • HIGH DUPLICATE/duplicate: Duplicate content found across components/cluster-azure/scripts/preflight.sh and components/inference-nim-operator/scripts/preflight.sh and components/osmo-azure/scripts/preflight.sh and components/osmo-k8s/scripts/preflight.sh: "kubectl_version()" in components/cluster-azure/scripts/preflight.sh (lines 139-145) vs "kubectl_semver()" in components/inference-nim-operator/scripts/preflight.sh (lines 60-65) vs "kubectl_semver()" in components/osmo-azure/scripts/preflight.sh (lines 123-128) vs "kubectl_semver()" in components/osmo-k8s/scripts/preflight.sh (lines 61-66) (components/cluster-azure/scripts/preflight.sh:139)

Publication Recommendation

The skill should be reviewed before NVSkills-Eval publication. Skill owners should address the findings above and rerun NVSkills-Eval to refresh this benchmark.

Source: SKILL.md on GitHub

2 warnings3mo3 checks · Risk SAFE
  • Gen Agent Trust Hub3mo

    This skill from NVIDIA provides a comprehensive set of tools and instructions for setting up Physical AI infrastructure on Azure AKS or local MicroK8s clusters. The analysis found no malicious patterns; all external downloads are from trusted or well-known services (NVIDIA, Microsoft, HashiCorp, Hugging Face), and the orchestration design prioritizes user visibility through a partitioned multi-agent architecture.

  • Socket3mo

    1 alert: gptAnomaly

  • Snyk3mo

    Risk: MEDIUM · 1 issue

Signed by skilld at 6c6fc09. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub yesterday.

Activeupdated 4 months ago
version
1.0.0
tools
[
  "Read",
  "Shell"
]
Other metadata
compatibility
Requires the selected component prerequisites, usually kubectl plus either MicroK8s or Azure CLI/Terraform, and OSMO or inference credentials for the chosen target.
metadata
{
  "author": "NVIDIA Physical AI",
  "tags": [
    "physical-ai",
    "infrastructure",
    "kubernetes",
    "azure",
    "microk8s",
    "osmo",
    "nim-operator",
    "scaling"
  ],
  "domain": "ai-ml",
  "languages": [
    "bash",
    "hcl",
    "yaml"
  ]
}

README badge

README badge for nvidia/skills/physical-ai-infrastructure-setup-and-resilient-scaling