Best Design & Media Claude SkillsPage 6

#SkillAuthorDownloadsStars
251
Audio Transcribe
Auto-transcribe voice messages locally using faster-whisper with selectable Whisper models, no API key required.
Alex Knight2.8K2
252
Venice AI Media
Generate, edit, and upscale images; create videos from images via Venice AI. Supports text-to-image, image-to-video (Sora, WAN), upscaling, and AI editing.
Neil2.8K6
253
Next Best Practices
Next.js best practices - file conventions, RSC boundaries, data patterns, async APIs, metadata, error handling, route handlers, image/font optimization, bund...
vi.dev2.8K2
254
Clips Machine
Transform long videos into viral short-form clips. Auto-detect best moments, add trendy captions, export for TikTok/Reels/Shorts. Self-contained, no external modules. 100% free tools.
Mayank82902.8K9
255
Image Process
Image processing tool for compression, background removal/replacement, and upscaling. Invoke when user wants to compress image, remove background, change bac...
Real2.8K-
256
Image Utils
Classic image manipulation with Python Pillow - resize, crop, composite, format conversion, watermarks, brightness/contrast adjustments, and web optimization...
Gal Davidi2.8K4
257
video-download
Download videos from 1800+ websites and generate subtitles using Faster Whisper AI. Use when user wants to download videos from YouTube, Bilibili, Twitter, T...
upupc2.8K3
258
Cine Cog
AI cinematic video production powered by CellCog. Short films, music videos, brand films, widescreen cinematics. Consistent characters, cinematic lighting, v...
CellCog2.8K2
259
YouTube Music
Manage YouTube Music library, playlists, and discovery via ytmusicapi.
gentrycopsy2.8K5
260
Video Understanding
视频理解与分析能力 - 让 AI 能够理解视频内容、提取关键信息。当用户要求分析视频、理解视频内容、总结视频、提取视频要点时触发此技能。
Jackeven022.8K1
261
Video Generator | 视频生成器
Automated text-to-video pipeline with multi-provider TTS/ASR support - OpenAI, Azure, Aliyun, Tencent | 多厂商 TTS/ASR 支持的自动化文本转视频系统
Justin Liu2.8K1
262
minimax-image
MiniMax AI图片生成,支持 image-01 和 image-01-live 模型。 image-01: 画面表现细腻,支持文生图、图生图 image-01-live: 手绘、卡通等画风增强,支持文生图并进行画风设置 需要 MiniMax API Key (Coding Plan)。
wanycun2.8K2
263
LiveAvatar
Talk face-to-face with your OpenClaw agent using a real-time video avatar powered by LiveAvatar
eNNNo2.7K9
264
SEO Content Engine
End-to-end SEO content creation workflow. Researches keywords, analyzes search intent, generates optimized outlines, and produces draft content with on-page...
Karl Ambrosius2.7K1
265
Photoshop Automator
Automate Adobe Photoshop on Windows via ExtendScript to run scripts, update text layers, create layers, apply filters, play actions, and export images.
abdul-karim-mia2.7K1
266
openai-tts-python
Text-to-speech conversion using OpenAI's TTS API for generating high-quality, natural-sounding audio. Supports 6 voices (alloy, echo, fable, onyx, nova, shimmer), speed control (0.25x-4.0x), HD quality model, multiple output formats (mp3, opus, aac, flac), and automatic text chunking for long content (4096 char limit per request). Use when: (1) User requests audio/voice output with triggers like "read this to me", "convert to audio", "generate speech", "text to speech", "tts", "narrate", "speak", or when keywords "openai tts", "voice", "podcast" appear. (2) Content needs to be spoken rather than read (multitasking, accessibility). (3) User wants specific voice preferences like "alloy", "echo", "fable", "onyx", "nova", "shimmer" or speed adjustments.
merend2.7K1
267
Azure Ai Voicelive Py
Build real-time voice AI applications using Azure AI Voice Live SDK (azure-ai-voicelive). Use this skill when creating Python applications that need real-time bidirectional audio communication with Azure AI, including voice assistants, voice-enabled chatbots, real-time speech-to-speech translation, voice-driven avatars, or any WebSocket-based audio streaming with AI models. Supports Server VAD (Voice Activity Detection), turn-based conversation, function calling, MCP tools, avatar integration, and transcription.
thegovind2.7K2
268
Inworld TTS
Text-to-speech via Inworld.ai API. Use when generating voice audio from text, creating spoken responses, or converting text to MP3/audio files. Supports multiple voices, speaking rates, and streaming for long text.
Gugic2.7K1
269
🩹 Image Inpainting — Pro Pack on RunComfy
Mask-driven image inpainting on RunComfy via the `runcomfy` CLI. Routes to Tongyi MAI Z-Image Turbo Inpainting (the dedicated inpainting endpoint with mask,...
Kalvin2.7K-
270
Coloring page
Turn an uploaded photo into a printable black-and-white coloring page.
Borahm2.7K1
271
BoTTube — AI Video Platform SDK
Browse, upload, and interact with videos on BoTTube (bottube.ai). Generate videos, prepare to constraints, upload, comment, and vote.
AutoJanitor2.7K5
272
B2C Mobile App Marketing Coach
The organic growth playbook behind 300K+ app downloads. Your AI becomes a growth coach trained on the exact system that drove 500M+ views and $30K+ revenue.
jackfriks2.7K16
273
Announcer
Announce text throughout the house via AirPlay speakers using Airfoil + ElevenLabs TTS.
Oliver Drobnik2.7K1
274
Local Whisper (cpp)
Local speech-to-text using whisper-cli (whisper.cpp).
wuxxin2.7K2
275
Wireframe
Create wireframes and user flows. Use when sketching page layouts in ASCII/SVG, mapping flows, or exporting to HTML.
BytesAgain22.7K-
276
💡 Relight — Pro Pack on RunComfy
Relight a still image — change the lighting setup, color temperature, direction, or mood — on RunComfy via the `runcomfy` CLI. Routes to Qwen Edit 2509's ded...
Kalvin2.7K-
277
Tube Cog
AI YouTube content creation powered by CellCog. YouTube videos, Shorts, thumbnails, video scripts, tutorials, vlogs, educational videos, product reviews, vid...
CellCog2.7K11
278
Sound FX
Generate short sound effects via ElevenLabs SFX (text-to-sound). Use when you need SFX clips like applause, canned laughter, whooshes, ambience, or short stingers, and optionally convert to WhatsApp-friendly .ogg/opus.
javicasper2.7K1
279
⏭️ Video Extend — Pro Pack on RunComfy
Extend or continue an existing video clip on RunComfy via the `runcomfy` CLI. Routes to Google Veo 3-1's `extend-video` and `fast/extend-video` endpoints — p...
Kalvin2.7K-
280
Podcast Generation from PDF, Text, and Links
Generate AI podcast episodes from PDFs, text, notes, and links using MagicPodcast in OpenClaw. Creates natural two-person dialogue audio, supports custom lan...
mogens92.7K4
281
🎞️ Video Inpainting — Pro Pack on RunComfy
Region edits across video frames on RunComfy via the `runcomfy` CLI — remove an object that appears across many frames, clean up wires or watermarks, replace...
Kalvin2.7K-
282
Fal Text-to-Image
Generate, remix, and edit images using fal.ai's AI models. Supports text-to-image generation, image-to-image remixing, and targeted inpainting/editing.
delorenj2.7K1
283
👄 Lipsync — Pro Pack on RunComfy
Lip-sync a face to a specific audio track on RunComfy via the `runcomfy` CLI. Routes across ByteDance OmniHuman (audio-driven full-body avatar from a portrai...
Kalvin2.7K-
284
🖼️ Image Outpainting — Pro Pack on RunComfy
Image outpainting on RunComfy via the `runcomfy` CLI — extend a still beyond its original canvas, fill in what the camera didn't capture, change aspect ratio...
Kalvin2.6K-
285
Flux Image
Generate images with FLUX models (Black Forest Labs) via inference.sh CLI. Models: FLUX Dev LoRA, FLUX.2 Klein LoRA with custom style adaptation. Capabilitie...
Ömer Karışman2.6K2
286
Generate images & videos with: Gemini 3 Pro Image (image) + Qwen Wan 2.6 (video) via one API key
Generate images & videos with AIsa. Gemini 3 Pro Image (image) + Qwen Wan 2.6 (video) via one API key.
0xjordansg-yolo2.6K12
287
3d Cog
AI 3D model generation powered by CellCog. Text-to-3D, image-to-3D — production-ready GLB files for games, AR/VR, e-commerce, and 3D printing. Game assets, p...
CellCog2.6K3
288
ElevenLabs
ElevenLabs API integration with managed authentication. AI-powered text-to-speech, voice cloning, sound effects, and audio processing. Use this skill when us...
byungkyu2.6K3
289
NK Images Search
Search 1+ million free high-quality AI stock photos. Generate up to 240 free AI images daily. No API key, no tokens, no cost. 235+ niches and growing.
tompltw2.6K10
290
🦴 ControlNet & Pose — Pro Pack on RunComfy
Pose-conditioned generation on RunComfy via the `runcomfy` CLI. Routes across Kling 2-6 Motion Control Pro / Standard (transfer the motion / blocking of a re...
Kalvin2.6K-
291
Google Photos Manager for OpenClaw
Manage Google Photos library. Upload photos, create albums, and list library content. Use when the user wants to backup, organize, or share images via Google Photos.
jorgermp2.6K3
292
Fal Ai
Generate images and media using fal.ai API (Flux, Gemini image, etc.). Use when asked to generate images, run AI image models, create visuals, or anything involving fal.ai. Handles queue-based requests with automatic polling.
Sxela2.6K3
293
MoodCast
Transform any text into emotionally expressive audio with ambient soundscapes using ElevenLabs v3 audio tags and Sound Effects API
Ashutosh Jha2.6K3
294
Roon Controller
Control Roon music player through Roon API with automatic Core discovery and zone filtering. Supports play/pause, next/previous track, and current track query. Automatically finds Muspi zones. Supports Chinese commands.
puterjam2.6K3
295
libtv-skills
agent-im 会话技能 - 通过 liblib.tv 的 AI 能力生成和编辑图片/视频。覆盖场景包括:生成(文生图、文生视频、图生视频、做动画、画一个xxx、来段xxx)、编辑修改(把xxx换成yyy、去掉xxx、加上xxx、改成xxx、调整xxx、局部修改、改镜头)、风格转换(风格迁移、转绘、换风格)、视...
3165307902.6K4
296
douyin-download
抖音无水印视频下载和文案提取工具
whille2.6K5
297
支持从 YouTube、Bilibili、抖音及所有 yt-dlp 兼容平台下载视频,可自动选择最佳分辨率、合并音视频并清理文件名
通用视频下载工具,支持 YouTube、B站、抖音等主流平台。使用 yt-dlp 下载视频,自动选择分辨率、合并音视频、清理文件名。
jingrongx2.6K8
298
Visla AI Video Creation
Creates AI-generated videos from text scripts, URLs, or PPT/PDF documents using Visla. Use when the user asks to generate a video, turn a webpage into a vide...
visla-admin2.6K-
299
fal
Search, explore, and run fal.ai generative AI models (image generation, video, audio, 3D). Use when user wants to generate images, videos, or other media with AI models.
apekshik2.6K1
300
📐 Video Outpainting — Pro Pack on RunComfy
Video outpainting on RunComfy via the `runcomfy` CLI — extend the spatial canvas of a video, change aspect ratio (9:16 vertical to 16:9 horizontal or vice ve...
Kalvin2.6K-
Browse all 4,490 Design & Media skills

More Categories

Category listings include Claude skills with 1,000+ downloads, merged from Clawhub & SkillHub. See methodology.