All skills
heygen-com avatar

/heygen-avatar

@1bd5e4d
by HeyGenheygen-com/skills460 stars
78

Create a persistent HeyGen avatar — a reusable face + voice identity for the agent, the user, or any named character — powered by HeyGen Avatar V technology. Prompt-based creation by default (description → HeyGen builds it); photo upload is optional for real-person digital twins. Use when: (1) giving the agent a face + voice so it can present videos ("bring yourself to life", "create your avatar", "give yourself an avatar", "design a presenter", "set up an avatar", "let's make an avatar"), (2) the user wants to appear in videos as themselves ("create my avatar", "I want my face in a video", "digital twin of me", "build me an avatar"), (3) building a named character presenter ("create an avatar called Cleo", "design a character named X"), (4) establishing HeyGen identity before making videos — the correct FIRST step when no avatar exists yet. Chain signal: when the user says both an identity/avatar action AND a video action in the same request ("create an avatar AND make a video", "set up identity THEN create a video", "design a presenter AND immediately record"), run heygen-avatar first, then heygen-video. Returns avatar_id + voice_id — pass directly to heygen-video to create HeyGen videos. NOT for: generating videos (use heygen-video), translating videos, or TTS-only tasks.

Use this Skill: https://skilld.dev/gh/heygen-com/skills/heygen-avatar

This session only. Nothing lands on disk.

referencesasset-routing.md

≈1.1k tokens on demand. Your agent reads this file only when SKILL.md points to it.

Asset Handling — The Classification Engine

When the user provides files, URLs, or references, route each asset to the right path. The user should NEVER have to think about this.

Two Paths

Path What happens When to use
A: Contextualize → Prompt Read/analyze the asset, extract key info, bake into script. Video Agent never sees the original. Reference material, auth-walled content, documents where the information matters more than the visual.
B: Attach to API Upload the raw file via files[]. Video Agent analyzes, extracts graphics, uses as frames/B-roll. Screenshots, branded assets, PDFs with important visual layouts, images the viewer should literally see.
A+B: Both Contextualize for script quality AND attach for visual use. Long docs where you need to summarize but Video Agent should also have the full source.

Classification Flow

1. Can Video Agent access this directly?
   - Public URL (no auth, no paywall) → YES
   - Private/internal URL → NO
   - Local file → NO (must upload first)

2. Should the viewer SEE this asset?
   - Screenshot, logo, product image, chart → YES → Path B
   - Research doc, article, context material → NO → Path A
   - Ambiguous → Path A+B

3. Is the content too long for the prompt?
   - Short (< 500 words) → fits in prompt
   - Long (> 500 words) → summarize key points, attach full doc

Decision Matrix

Asset Type Publicly Accessible? Show On Screen? Route
Screenshot / image N/A Yes B: Attach + describe in prompt as B-roll
Logo / brand asset N/A Yes B: Attach + anchor to intro/outro
Public URL to file (PDF, image, video) Yes Maybe B: Download → upload via /v3/assets → pass asset_id + summarize
Public URL to web page (HTML) Yes No A: Fetch and contextualize only. Do NOT pass HTML URLs in files[].
Auth-walled URL (requires login) No No A: Ask the user to paste the content. Never fabricate.
PDF (short, text-heavy) N/A No A+B: Extract key points + attach
PDF (long, visual-rich) N/A Maybe B: Attach + summarize top points
Raw data / spreadsheet N/A Partially A: Analyze and describe key stats. Attach if charts should appear.

Executing Routes

Path A (Contextualize)

  • URLs: Use web_fetch to retrieve publicly accessible content
  • For auth-walled content you cannot access: ask the user to paste the text directly
  • Extract 3-5 most important points relevant to the video
  • Weave naturally into the script. Don't dump. Integrate.

Path B (Attach)

Upload to HeyGen:

MCP: upload via the asset tool (depends on environment). CLI: heygen asset create --file /path/to/file.png

Max 32MB per file. Returns JSON with the new asset_id.

Or pass inline in files[]:

{"type": "url", "url": "https://example.com/image.png"}
{"type": "asset_id", "asset_id": "<from upload>"}
{"type": "base64", "data": "<base64>", "media_type": "image/png"}

Describe Asset Usage in Prompt

Be SPECIFIC:

  • "Use the uploaded dashboard screenshot as B-roll when discussing analytics"
  • "Display the company logo in the intro and end card"

Log Classification

In the learning log entry, record:

"assets_classified": [{"type": "image", "route": "attach", "accessible": true, "reason": "product screenshot"}]

Rules

  • Never ask the user which path unless genuinely 50/50. You're the producer. Make the call.
  • When in doubt, do both (A+B). Over-providing costs nothing.
  • Always describe attached assets in the prompt. Uploading without description = ignored.
  • Auth-walled content is YOUR job. Bridge the gap between your access and Video Agent's.
  • URLs that fail: Try web_fetch. If login/paywall/404 → tell the user, ask for content directly. Never silently fabricate.
  • HTML URLs cannot go in files[]. Video Agent rejects text/html. Web pages are ALWAYS Path A only.
  • Prefer download→upload→asset_id over files[]{url}. HeyGen's servers often blocked by CDN/WAF.

Source: SKILL.md on GitHub

No alerts16d3 checks · Risk SAFE
  • Gen Agent Trust Hub16d

    The skill is safe. It enables the creation of digital avatars using the HeyGen platform and includes vendor-provided instructions for setting up the HeyGen CLI. While it processes user-provided identity data, it employs structured mappings to interface with the HeyGen API.

  • Socket16d

    No alerts

  • Snyk16d

    Risk: LOW · No issues

Signed by skilld at 1bd5e4d. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub last week.

Steadyupdated 3 months ago
What it can do
Runs commands Network Reads files Edits files
MCP servers
heygen
version
3.2.0
argument-hint
[name_or_description]
All 5 allowed tools
BashWebFetchReadWritemcp__heygen__*

README badge

README badge for heygen-com/skills/heygen-avatar