SKILL.md packages that extend Claude Code, Cursor, Copilot, and other AI agents.
Tags

video-podcast-maker
Automates end-to-end production of long-form video podcasts using TTS, Remotion, and FFmpeg—supports multi-language output and Bilibili/YouTube publishing workf

voxclaw
A macOS menu-bar app that lets agents send text to a local Mac for speech (Apple TTS or OpenAI voices) over HTTP.

skills
Guides server-side integration of Runway audio APIs: TTS, sound effects, voice isolation, dubbing and speech-to-speech conversion.

qwen3_tts_rs
Generate speech audio from text or clone voices from reference audio using Qwen3 TTS binaries and models; supports multiple named speakers, English and Chinese.

trending-skills
Generate interactive AI-powered lessons (slides, quizzes, simulations) from any topic or document using a multi-agent pipeline and TTS-enabled agents.

cc-skills
Start, stop, and verify a local Kokoro TTS HTTP server (OpenAI-compatible /v1/audio/speech) with health checks and troubleshooting guidance.

video-zebra-china
Download and localize foreign-language videos into Simplified Chinese: transcripts, translated subtitles, Mandarin TTS dubbing, audio mix, and final MP4 export.

babysor
Convert text (or SRT timelines) into speech audio using local Kokoro or Noiz cloud backends, with voice cloning and timeline-aligned rendering.

xiaotianfotos
Multi-engine AI creative toolchain: TTS, ASR, image generation (ComfyUI), and subtitle/video cut tooling.

code-explainer
Interactive code walkthroughs that scan a codebase, plan segments, and narrate highlights with configurable depth and VS Code integration.

narrator-ai-cli-skill
CLI orchestration skill to generate narrated videos (movie commentary, short-drama dubbing) using Narrator AI: material selection, BGM, voice dubbing, writing a

krillinai
Plan and run multi-stage video localization workflows (subtitle, TTS, render) for translation and dubbing across formats and languages.

device-takeover
Control Android and Linux devices over the network: capture screen, send touch/keyboard input, run commands, and perform voice I/O for remote control or gamepla

skills
Create podcasts, explainer videos, TTS, and AI images using ListenHub scripts; run the provided shell scripts to generate, check status, and download outputs.

aaas-vault
Create calming, slow-paced sleep stories and templates for bedtime audio: sensory-rich, low-stakes narratives designed to help listeners drift off.

claude-skill-registry
Enables natural voice conversations for AI assistants using local (Whisper, Kokoro) or cloud-based Speech-to-Text (STT) and Text-to-Speech (TTS) services.

ankitjh4
Integration for Sarvam AI: TTS (Bulbul), STT (Saarika), translation, transliteration, and document intelligence with examples and best practices.

skills
Generate TTS, sound effects, voice isolation, dubbing, and speech-to-speech conversions via Runway's API using provided runnable scripts.

sutando
Build a concise, shareable 30–60s news-explainer video with validated assets, TTS, and an automated pre-publish gate.

claude-code-skills
Cost-effective AI media tools for transcription, image generation, TTS, OCR, and video creation via the deAPI decentralized network.

awesome-omni-skill
Automates the conversion of PPTX files or structured JSON slides into professional MP4 videos with AI narration and subtitles.