noizai/skills▌
6 approved skills in this repository
template-skill
Productivity
$21
video-translation
Video
Translate video speech into another language with AI-generated dubbing that preserves original timing and emotion. \n \n Downloads videos and subtitles, translates subtitle text, then generates dubbed audio using TTS with voice cloning matched to the original speaker's tone \n Automatically aligns dubbed audio duration to original subtitle timestamps and preserves background audio outside speech segments \n Requires youtube-downloader and tts skills as dependencies, plus ffmpeg and a Noiz API ke
chat-with-anyone
Productivity
Clone real voices from online video or design voices from photos, then roleplay as that person with synthetic speech. \n \n Two workflows: extract voice from public video (interviews, speeches) by name, or generate a matching voice from an uploaded image of an unrecognizable person \n Requires ffmpeg , yt-dlp , the tts skill, and a Noiz API key; includes setup verification and dependency installation steps \n Built-in ethical guardrails: agent must refuse requests targeting non-consenting privat
tts
Productivity
Text-to-speech with dual backends, voice cloning, and timeline-accurate audio synthesis for dubbing and video narration. \n \n Supports two backends: Kokoro (local, offline) for simple speech synthesis, and Noiz (cloud) for voice cloning, emotion control, and precise segment timing \n Simple mode converts text, files, or URLs to audio with optional voice cloning from reference audio; timeline mode aligns speech to SRT subtitles with per-segment voice and emotion control \n Voice maps enable gran
daily-news-caster
AI/ML
Fetches latest news, converts it to a dual-host podcast script, and generates audio using text-to-speech. \n \n Requires news-aggregator-skill and tts skill as dependencies; installation commands provided if not present \n Generates conversational Q&A-style podcast scripts in Markdown with two hosts asking and answering questions about news items \n Produces audio line-by-line using the tts skill, then concatenates with ffmpeg into a single podcast file \n Supports reference audio files for
characteristic-voice
Productivity
Add human warmth, emotion, and natural speech patterns to AI voice output. \n \n Includes five speaking presets (good night, good morning, comfort, celebration, just chatting) that automatically tune pace, warmth, and emotional tone \n Sprinkle non-lexical fillers (hmm, aww, haha, sighs) at natural pauses to create conversational authenticity; supports up to 4 fillers per short message \n Clone specific character voices by providing reference audio from YouTube clips or personal recordings, forw