heygen-video
Generate HeyGen presenter videos via the v3 Video Agent pipeline — handles Frame Check (aspect ratio correction), prompt engineering, avatar resolution, and voice selection. Required for any HeyGen video generation. Replaces deprecated endpoints with v3. Use when: (1) generating any HeyGen video (via API or otherwise), (2) sending a personalized video message (outreach, update, announcement, pitch, knowledge), (3) creating a HeyGen presenter-led explainer, tutorial, or product demo with a human
Security Assessment
About heygen-video
This skill generates HeyGen presenter (talking-head) videos through HeyGen's v3 Video Agent pipeline. It solves the problem of producing high-quality AI presenter videos by orchestrating the steps that raw API calls skip: aspect-ratio and framing correction ("Frame Check"), prompt engineering, avatar resolution, voice selection, and asset routing. It explicitly forbids calling deprecated v1/v2 endpoints directly, routing instead through MCP, the OpenClaw plugin, or the `heygen` CLI, and returns a shareable video URL plus a session URL for iteration.
The skill acts as a guided "video producer." It reads saved avatar identity files (AVATAR-<NAME>.md, plus role symlinks) to load an avatar's group and voice IDs, appends framing-correction notes to the prompt based on avatar orientation and background type, and classifies user-supplied assets into two paths — contextualize the content into the script, or upload the raw file to HeyGen for on-screen B-roll. It downloads voice preview audio to a temporary directory for playback, appends one JSON line per generated video to a local learning log, and follows strict UX rules (be concise, hide internal pipeline jargon, poll silently for completion). A companion `update-check.sh` script is explicitly opt-in and is not run automatically on invocation. Authentication uses a HEYGEN_API_KEY environment variable.
Target users are people creating personalized video messages, outreach, announcements, pitches, explainers, tutorials, or product demos with a human presenter. When a request also involves creating or designing an avatar, the skill chains to the companion heygen-avatar skill first. It is not for avatar creation, cinematic b-roll without a presenter, video translation, TTS-only output, or streaming avatars.
FAQ
What do I need to use this skill?
A HeyGen API key supplied via the HEYGEN_API_KEY environment variable, and access to HeyGen via MCP, the OpenClaw plugin, or the heygen CLI. It only calls the v3 pipeline and refuses deprecated v1/v2 endpoints.
Can I use my own avatar?
Yes. It accepts an avatar_id (for example from the heygen-avatar skill) and reads saved AVATAR-<NAME>.md identity files for the avatar's group and voice IDs, or falls back to a stock presenter. If you provide a photo and want a video, it routes to heygen-avatar first.
What is Frame Check?
An internal step that runs when an avatar_id is set: it fetches avatar look metadata, determines orientation and background, and appends framing-correction notes to the prompt so the presenter is composed correctly for landscape, portrait, or square-to-landscape videos. It does not generate images or new looks.
How does it handle files and URLs I provide?
It classifies each asset into contextualizing it into the script, uploading it to HeyGen (max 32MB per file) for on-screen B-roll, or both. HTML web pages are only fetched for context and never attached, and auth-walled content is requested from the user rather than fabricated.
What is the update-check script and does it run automatically?
It is an opt-in scripts/update-check.sh that checks for skill updates. The skill states it must not be executed automatically on invocation; you run it manually only when desired.
All Files
10 filesInstall heygen-video
Quick Setup:
- Copy the skill folder to
.claude/skills/ - Claude will automatically detect and use the skill
Repository
heygen-com/skills