heygen-com/skills▌
11 approved skills in this repository
video-understand
Video
Understand video content locally using ffmpeg for frame extraction and Whisper for transcription. Fully offline, no API keys required.
ai-video-gen
AI/ML
Generate AI videos from text prompts. Supports multiple providers (VEO 3.1, Kling, Sora, Runway, Seedance), configurable aspect ratios, and optional reference images for image-to-video generation.
text-to-speech
Productivity
Generate speech audio from text using HeyGen's Starfish TTS model with voice, speed, and pitch control. \n \n List available TTS voices by language and gender, then generate audio files with customizable speed (0.5–1.5) and pitch (−50 to 50) \n Supports multilingual voices with locale selection (e.g., pt-BR ) and SSML-style break tags for pauses within text \n Returns audio URL, duration, request ID, and word-level timestamps for caption syncing or timed overlays \n Requires HEYGEN_API_KEY envir
heygen
Productivity
[DEPRECATED] Legacy avatar video generation API — use create-video or avatar-video skills instead. \n \n Generates talking-head videos, explainers, and presentations from text prompts or detailed scripts with precise avatar, voice, and scene control \n Provides MCP tools for prompt-based generation, video status polling, and account management; falls back to direct HTTP API calls if tools unavailable \n Supports multi-scene videos with per-scene avatars, backgrounds, voices, and timing; captions
faceswap
Productivity
Swap a face from a source image into a target video using GPU-accelerated AI processing. The source image provides the face to swap in, and the target video receives the new face.
visual-style
Productivity
Create, extract, and apply portable visual design systems. A visual-style.md file defines colors, typography, layout, motion, and mood in one file that any AI tool can consume.
video-download
Video
Download video and audio from URLs using yt-dlp directly. No wrapper scripts needed.
create-video
Video
Generate complete videos from a text prompt. Describe what you want and the AI handles script writing, avatar selection, visuals, voiceover, pacing, and captions automatically.
video-edit
Video
Edit videos locally by running ffmpeg/ffprobe directly. No wrapper scripts needed.
video-translate
Video
Translate and dub videos into multiple languages with automatic lip-sync using HeyGen. \n \n Supports 12+ languages including Spanish, French, German, Japanese, Chinese, and Arabic with automatic voice cloning to match original speaker characteristics \n Offers two translation modes: full lip-sync dubbing or faster audio-only translation without video adjustments \n Handles multi-speaker videos, custom vocabulary preservation, and optional background music removal or speech enhancement \n Includ
avatar-video
Video
Create AI avatar videos with full control over avatars, voices, scripts, and backgrounds using POST /v3/videos. Two creation modes via discriminated union on type: