Best Design & Media Claude SkillsPage 5

#SkillAuthorDownloadsStars
201
🫧 Video Edit — Pro Pack on RunComfy
Video edit on RunComfy. This video edit skill transforms an existing video clip — restyle, background swap, outfit swap, motion transfer, color grade, or any...
Kalvin3.2K2
202
Youtube Editor
Automate YouTube video editing: download videos, transcribe with Whisper, analyze content using GPT-4, and create Korean SEO-optimized metadata plus consiste...
jeong-wooseok3.2K-
203
AI Notes Video
Generate detailed AI notes including document, outline, and image-text formats from a user-provided video URL using Baidu's video AI notes tool.
baidu_qianfan3.2K2
204
Sketch Illustration
插画图片生成技能,支持多种手绘风格。适合流程图、功能说明、PPT配图、教程配图、知识图和手绘信息图等场景。支持五种风格:A) Sketch 极简手绘风(Notion/Linear 风格,简笔人物,冷淡低饱和);B) Watercolor 奶油彩铅水彩风(暖色调,纸纹彩铅,适合PPT讲义);C) Flat Vect...
dadaniya993.2K1
205
🫧 Image-to-Video — Pro Pack on RunComfy
Image-to-video generation on RunComfy. This image-to-video skill turns any still image into a short video clip via the RunComfy Model API. The image-to-video...
Kalvin3.2K1
206
Elevenlabs Transcribe
Transcribe audio to text using ElevenLabs Scribe. Supports batch transcription, realtime streaming from URLs, microphone input, and local files.
PaulAsjes3.2K2
207
Nanobanana Pro
Nano Banana Pro with auto model fallback — generate/edit images via Gemini Image API. Run via: uv run {baseDir}/scripts/generate_image.py --prompt 'desc' --filename 'out.png' [--resolution 1K|2K|4K] [-i input.png]. Supports text-to-image + image-to-image (up to 14); 1K/2K/4K. Fallback chain: gemini-2.5-flash-image → gemini-2.0-flash-exp. MUST use uv run, not python3.
yazelin3.2K4
208
Ai Video Generation
Generate AI videos with Google Veo, Seedance, Wan, Grok and 40+ models via inference.sh CLI. Models: Veo 3.1, Veo 3, Seedance 1.5 Pro, Wan 2.5, Grok Imagine...
Ömer Karışman3.1K2
209
Video Translator
Real time video translation / dubbing skill. Translate user-provided video (file or URL) and return preview_url. 适用于视频直译、视频翻译、视频配音、字幕翻译出片。
papawaigo3.1K6
210
douyin-video
抖音视频下载工具 - 解析抖音链接,下载视频并发送
andjie983.1K-
211
Comfyui-Api
Connects to a ComfyUI server to generate images from prompts, auto-detects URLs, translates Chinese prompts, and supports REST and WebSocket APIs.
qqliaoxin3.1K1
212
🧰 RunComfy CLI — Pro Pack on RunComfy
RunComfy CLI on RunComfy. The `runcomfy` CLI is one binary, one auth, hundreds of RunComfy model endpoints — image generation on RunComfy, image edit on RunC...
Kalvin3.1K-
213
feishu-send-file
飞书发送文件技能。用于通过飞书向用户发送普通文件附件(HTML、ZIP、PDF、代码文件等)以及处理“本地图片路径被发成路径文本”的可靠补救场景。普通文件必须先上传获取 `file_key` 再发送;当本地图片用 `message`/`media` 发送后在飞书里只显示 `/root/...png` 路径而不显示...
dadaniya993.1K3
214
Podcast Generation with Microsoft Foundry
Generate AI-powered podcast-style audio narratives using Azure OpenAI's GPT Realtime Mini model via WebSocket. Use when building text-to-speech features, audio narrative generation, podcast creation from content, or integrating with Azure OpenAI Realtime API for real audio output. Covers full-stack implementation from React frontend to Python FastAPI backend with WebSocket streaming.
thegovind3.1K3
215
🎬 AI Video Generation — Pro Pack on RunComfy
AI video generation on RunComfy. This RunComfy video generation skill is a smart router across the RunComfy video-model catalog — HappyHorse 1.0 (Arena #1, n...
Kalvin3.1K-
216
Design System Creation
Complete workflow for creating distinctive design systems from scratch. Orchestrates aesthetic documentation, token architecture, components, and motion. Use when starting a new design system or refactoring an existing one. Triggers on create design system, design tokens, design system setup, visual identity, theming.
wpank3.1K6
217
it will help you to send voice messages to your AI Assistant and also can make it talk
Text-to-Speech and Speech-to-Text using ElevenLabs AI. Use when the user wants to convert text to speech, transcribe voice messages, or work with voice in multiple languages. Supports high-quality AI voices and accurate transcription.
amreahmed3.1K1
218
Nano Banana 2
Generate images with Google Gemini 3.1 Flash Image Preview (Nano Banana 2) via inference.sh CLI. Capabilities: text-to-image, image editing, multi-image inpu...
Ömer Karışman3.1K5
219
Media Writing
You are a professional media writing expert with extensive experience in creating engaging and impactful content across multiple formats. Creating attention-...
alvinecarn3.1K1
220
🎨 AI Image Generation — Pro Pack on RunComfy
AI image generation on RunComfy. This RunComfy image generation skill is a smart router across the RunComfy image-model catalog — FLUX 2 (Klein 9B/4B, Pro, D...
Kalvin3.1K-
221
Jianying
提供剪映官网公开的模板分类、编辑功能、导出规格、版权说明及套餐权益等信息整理和说明服务。
codenova583.1K1
222
Donson Intelligent Editing
Use when performing video/audio processing tasks including transcoding, filtering, streaming, metadata manipulation, or complex filtergraph operations with FFmpeg.
DonsonAI3.1K3
223
Mia Content Creator
AI agent content creation and monetization across platforms
ArubikU3.0K3
224
Walkie-Talkie Mode
Handles voice-to-voice conversations on WhatsApp. Automatically transcribes incoming audio and responds with local TTS audio. Use when the user wants to "talk" instead of type.
Rubén Fernández Boullón3.0K4
225
Cad
cad reference tool
bytesagain-lab3.0K-
226
Yollomi AI Image & Video Generator
AI image generator skill (image, image generation). Multi-model image generator for Yollomi to generate AI images via one unified API endpoint. Requires YOLL...
AniChikage3.0K6
227
clawtunes
Control Apple Music on macOS via the `clawtunes` CLI (play songs/albums/playlists, control playback, volume, shuffle, repeat, search, catalog lookup, AirPlay...
forketyfork3.0K3
228
Seedream5
使用火山引擎豆包 Seedream 5.0 生成或编辑图片。支持文生图、单图/多图生图、组图、联网搜索。
zealman20253.0K-
229
产品prd转设计文档
将产品需求文档(PRD)转换为完整的设计需求文档,包含信息架构、交互流程、页面布局、视觉规范等内容,并自动生成交互流程图。适用于产品设计、交互设计、UI设计等场景。当用户提供PRD文档并需要设计需求文档时触发此技能。
user_8f3f44073.0K17
230
UI设计Skill
专业UI生成与现有界面优化技能。当用户请求生成UI界面、设计页面、优化现有界面、创建组件库、设计系统、App界面、Web界面、Dashboard、Landing Page、或任何涉及视觉交互设计时必须触发此技能。参考Google Material Design、Apple HIG、以及Pinterest级别审美标准,输出大气简约、高级美学的UI代码。即使用户只是说"帮我做个界面"、"优化一下UI"、"设计一个页面"也必须触发。
user_865e35923.0K14
231
Aliyun Asr
Pure Aliyun ASR skill for voice message transcription, supports multiple channels including Feishu
Jixson3.0K2
232
ClawHub - YouTube Downloader & Clipper
Clip and download specific time ranges or full YouTube videos in various qualities, including audio-only MP3 extraction, using precise timestamps.
sandeepyadav14782.9K1
233
Image Generate
使用内置 image_generate.py 脚本生成图片, 准备清晰具体的 `prompt`。
zhangsiyuan1231232.9K-
234
Meitu Skills
Comprehensive Meitu AI toolkit for image and video editing. Features include AI poster design, precise background cutout, virtual try-on, e-commerce product...
Meitu.Inc2.9K121
235
Media Player
Play audio/video locally on the host
Xejrax2.9K3
236
Diagram
Generate diagrams from descriptions with Mermaid, PlantUML, or ASCII for architecture, flows, sequences, and data models.
Iván2.9K3
237
Runware Image & Video generation
Generate images and videos via Runware API. Access to FLUX, Stable Diffusion, Kling AI, and other top models. Supports text-to-image, image-to-image, upscaling, text-to-video, and image-to-video. Use when generating images, creating videos from prompts or images, upscaling images, or doing AI image transformation.
26medias2.9K2
238
ardot-showcase-autolayout
用于在 Ardot 中把多个已有画板自动编排成可交付的衍生展示区。适合多状态对照、设计评审看板、流程演示、角色/设备矩阵。也可作为 CodeBuddy + Ardot AI 设计工作流的收尾节点,将 AI 生成的零散画板自动整理成结构化展示区并输出可分享链接。
user_da6c07642.9K9
239
Home Music
Control whole-house music scenes combining Spotify playback with Airfoil speaker routing. Quick presets for morning, party, chill modes.
Andy Steinberger2.9K1
240
GifHorse
Search video dialogue and create reaction GIFs with timed subtitles. Perfect for creating meme-worthy clips from movies and TV shows.
Coyote-git2.9K1
241
Yt Dlp
A robust CLI wrapper for yt-dlp to download videos, playlists, and audio from YouTube and thousands of other sites. Supports format selection, quality control, metadata embedding, and cookie authentication.
azzar budiyanto2.9K-
242
Whisper Transcribe
Transcribe audio files to text using OpenAI Whisper. Supports speech-to-text with auto language detection, multiple output formats (txt, srt, vtt, json), batch processing, and model selection (tiny to large). Use when transcribing audio recordings, podcasts, voice messages, lectures, meetings, or any audio/video file to text. Handles mp3, wav, m4a, ogg, flac, webm, opus, aac formats.
Jonas Pfalzgraf2.9K5
243
fame graphic
Generate diverse creative illustrations via OpenAI Images API. Create book illustrations, editorial art, children's book art, concept illustrations, and artistic scenes. Use when user needs creative visual content for stories, articles, presentations, or artistic projects (e.g., "illustrate a fairy tale scene", "create editorial art about technology", "design children's book illustrations", "generate concept art for a story").
adebayoabdushaheed-a11y2.9K3
244
Elite Frontend Design
前端 UI 界面设计。当用户要创建网页、landing page、dashboard、React/Vue 组件、前端页面时触发。 输出 HTML/CSS/JS 代码。不适用于:静态图片设计(用 canvas-design)、公众号配图(用 weixin-canvas-design)。
莫循2.9K1
245
seedance-2-video-gen
Seedance 2.0 AI video generation via EvoLink API. Three modes — text-to-video, image-to-video (1-2 images), reference-to-video (images + videos + audio). Aut...
EvolinkAI2.9K6
246
Parakeet Stt
Local speech-to-text with NVIDIA Parakeet TDT 0.6B v3 (ONNX on CPU). 30x faster than Whisper, 25 languages, auto-detection, OpenAI-compatible API. Use when transcribing audio files, converting speech to text, or processing voice recordings locally without cloud APIs.
carlulsoe2.9K1
247
Lead Magnets
Design specific, quick-win lead magnets that attract qualified prospects, capture emails, and smoothly convert them into paying customers.
staybased2.8K3
248
Airfoil
Control AirPlay speakers via Airfoil from the command line. Connect, disconnect, set volume, and manage multi-room audio with simple CLI commands.
Andy Steinberger2.8K-
249
Reve AI Image Generation
Generate, edit, and remix images using the Reve AI API. Use when creating images from text prompts, editing existing images with instructions, or combining/remixing multiple reference images. Requires REVE_API_KEY or REVE_AI_API_KEY environment variable.
dpaluy2.8K1
250
macOS Local Voice
Local STT and TTS on macOS using native Apple capabilities. Speech-to-text via yap (Apple Speech.framework), text-to-speech via say + ffmpeg. Fully offline, no API keys required. Includes voice quality detection and smart voice selection.
STRRL2.8K1
Browse all 4,490 Design & Media skills

More Categories

Category listings include Claude skills with 1,000+ downloads, merged from Clawhub & SkillHub. See methodology.