Best Design & Media Claude SkillsPage 2

#SkillAuthorDownloadsStars
51
Ai Image Generation
Generate AI images with FLUX, Gemini, Grok, Seedream, Reve and 50+ models via inference.sh CLI. Models: FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image,...
Ömer Karışman8.5K3
52
AI Image Gen
AI图像生成与编辑。支持文生图、图+文生图、风格转换。当用户要求画图、生成图片、编辑图片、图片风格转换时使用此 skill。支持多种比例(1:1、3:2、16:9、21:9 等)和分辨率(标准、2K、4K)。
Nick Qiu8.0K2
53
Graphic Design
Support design understanding from basic visuals to professional production and theory.
Iván7.8K17
54
Faster Whisper
Local speech-to-text using faster-whisper. 4-6x faster than OpenAI Whisper with identical accuracy; GPU acceleration enables ~20x realtime transcription. SRT...
Sarah Mak7.7K10
55
FFmpeg
Process video and audio with correct codec selection, filtering, and encoding settings.
Iván7.6K8
56
Video Agent (Deprecated)
[DEPRECATED — uses outdated v1/v2 endpoints] Use `create-video` for prompt-based video generation (v3 Video Agent) or `avatar-video` for precise avatar/scene...
Michael Wang7.6K7
57
Elevenlabs Tts
ElevenLabs TTS - the best ElevenLabs integration for OpenClaw. ElevenLabs Text-to-Speech with emotional audio tags, ElevenLabs voice synthesis for WhatsApp,...
shaharsh7.5K7
58
Image Vision
Analyze and interpret images by describing content, extracting text, answering questions, comparing visuals, and extracting structured data from JPG, PNG, GI...
cntuang7.4K4
59
OpenAI TTS
Text-to-speech via OpenAI Audio Speech API.
Mark Pors 🦖7.4K6
60
seedance2-skill
即梦 Seedance 视频创意工作台。用户发图+文案时自主完成看图分析→文案扩写→运镜匹配→质量验证→API生成。触发词:即梦、Seedance、seedance、视频生成、视频提示词、AI视频、运镜、短剧、广告视频、视频延长、图生视频。
zhanghaonan7777.4K16
61
Audio Cog
AI audio generation and text-to-speech powered by CellCog. Voiceover, narration, voice cloning, avatar voices, sound effects, music, podcasts, dialogue. Thre...
CellCog7.4K9
62
Voice Transcribe
Transcribe audio files using OpenAI's gpt-4o-mini-transcribe model with vocabulary hints and text replacements. Requires uv (https://docs.astral.sh/uv/).
darinkishore7.3K13
63
Seedream 图片生成
Seedream 图片生成 - 火山引擎方舟大模型服务平台图片生成 API。支持文生图、图生图、多图融合、组图生成等多种模式。
Xiao Huang7.3K5
64
ElevenLabs Voices
High-quality voice synthesis with 18 personas, 32 languages, sound effects, batch processing, and voice design using ElevenLabs API.
Robby7.3K17
65
Jimeng AI
基于火山引擎即梦AI的文生图/文生视频能力,支持通过文本描述生成图片和视频。
ogenes7.2K4
66
Kokoro TTS
Generate spoken audio from text using the local Kokoro TTS engine. Use when the user asks to "say" something, requests a voice message, or wants text converted to speech.
edkief7.2K1
67
视频生成
使用火山引擎 SD1.5pro API 生成视频。支持文本到视频和图生视频,异步处理任务。
user_cb5e28417.1K15
68
Mac TTS
Text-to-speech using macOS built-in `say` command. Use for voice notifications, audio alerts, reading text aloud, or announcing messages through Mac speakers. Supports multiple languages including Chinese (Mandarin), English, Japanese, etc.
kalijason7.0K3
69
Nano Banana Pro Image Gen(基于API易代理站)
图片生成技能,当用户需要生成图片、视觉信息图、创建图像、编辑/修改/调整已有图片时使用此技能。基于中国的API易代理站(https://api.apiyi.com/)的NanoBananaPro模型的图片生成服务,无需访问外网。支持10种宽高比的图片比例(`1:1`、`16:9`、`9:16`、`4:3`、`3:...
无处不在6.7K3
70
AI Full media generation automation flow Skill
VAP Media API skill for image, video, music, and media editing through VAP. Uses VAP product keys with current Media API endpoints.
VAP Media automation flow6.7K13
71
Video
Process, edit, and optimize videos for any platform with compression, format conversion, captioning, and repurposing workflows.
Iván6.5K9
72
FFmpeg CLI
Process video and audio using FFmpeg CLI for transcoding, cutting, merging, audio extraction, thumbnails, GIFs, speed, filters, subtitles, and watermarks.
ascendswang6.5K13
73
Veo
Generate video using Google Veo (Veo 3.1 / Veo 3.0).
Buddy Hadry6.4K4
74
LiblibAI Image & Video Gen
Generate images with Seedream4.5 and videos with Kling via LiblibAI API. Use when user asks to generate/create images, pictures, illustrations, or videos using LiblibAI, Seedream, or Kling models.
远方青木6.2K1
75
Tailwind Design System
Build scalable, themable Tailwind CSS component libraries using CVA for variants, compound components, design tokens, dark mode, and responsive grids.
wpank6.2K8
76
Douyin Cover Builder(抖音封面生成器)
这是一个面向中文创作者的 OpenClaw Skill,输入主题与人物气质后,会输出可直接用于生图模型的高质量提示词与创意说明。
0x000000036.2K9
77
ComfyUI
Run local ComfyUI workflows via the HTTP API. Use when the user asks to run ComfyUI, execute a workflow by file path/name, or supply raw API-format JSON; supports the default workflow bundled in assets.
kelvincai5226.2K3
78
图片提示词生成
基于五层拆解法的AI图片提示词生成器。将模糊的创意想法转化为结构严谨、可执行的图像生成规格书,支持多种风格预设和目标工具适配。
349840432m-dev6.1K9
79
Canva
Create, export, and manage Canva designs via the Connect API. Generate social posts, carousels, and graphics programmatically.
abgohel6.1K11
80
UI UX Pro Max
Professional UI/UX design resource library with searchable design patterns, color palettes, font pairings, chart types, and UX guidelines. Use when creating...
iReact2Code6.1K16
81
Google Gemini Media
Use the Gemini API (Nano Banana image generation, Veo video, Gemini TTS speech and audio understanding) to deliver end-to-end multimodal media workflows and code templates for "generation + understanding".
Xsir06.0K6
82
Tencent MPS
腾讯云 MPS 媒体处理服务,凡涉及以下任一场景必须触发:【视频转码】转码/压缩/格式转换/H.264/H.265/AV1/MP4/编码/码率/分辨率/帧率。【画质增强】画质增强/老片修复/超分/视频超分/真人增强/漫剧增强/防抖/720P/1080P/2K/4K。【音频处理】音频分离/人声提取/伴奏提取/去人声...
mpaas-skills6.0K6
83
Jimeng Video Generator
即梦AI视频生成工具(带声音版本),通过火山引擎API自动生成带音频的高质量视频。支持文生视频、图生视频,适用于短视频内容创作。
Buck5.9K5
84
Gemini Image Gen
Generate and edit images via Google Gemini API. Supports Gemini native generation, Imagen 3, style presets, and batch generation with HTML gallery. Zero depe...
김덕환5.9K9
85
SVG Draw
Create SVG images and convert them to PNG without external graphics libraries. Use when you need to generate custom illustrations, avatars, or artwork (e.g., "draw a dragon", "create an avatar", "make a logo") or convert SVG files to PNG format. This skill works by writing SVG text directly (no PIL/ImageMagick required) and uses system rsvg-convert for PNG conversion.
LiJY20155.7K7
86
Douyin Video Fetch
下载抖音视频到本地(无水印优先)。用于给后续视频分析/复刻提供原始素材,支持 URL 或 video_id 输入、批量列表输入与统一输出目录。
KAMIENDER5.7K20
87
Nano Banana Pro
Generate or edit images via Gemini 3 Pro Image (Nano Banana Pro).
DyCathecorde5.7K19
88
Computer Vision Expert
SOTA Computer Vision Expert (2026). Specialized in YOLO26, Segment Anything 3 (SAM 3), Vision Language Models, and real-time spatial analysis.
Zorro Ng5.6K4
89
图像生成 / Image Generation
Image Generation via Coze | 基于 Coze 的图像生成技能 Generate images using Coze workflows. 使用 Seedream 4.5 model. Handles parameter building and result parsing. 负责参数构...
noah5.6K-
90
speech-recognition
通用语音识别 Skill。支持多种音频格式(ogg/mp3/wav/m4a),使用硅基流动 SenseVoice API 进行语音转文字。当用户发送语音消息、音频文件,或需要转录音频时触发。
demo1125.5K3
91
Elevenlabs
Text-to-speech, sound effects, music generation, voice management, and quota checks via the ElevenLabs API. Use when generating audio with ElevenLabs or mana...
Oliver Drobnik5.5K2
92
Wan Image and Video Generation and Editting
Image and Video Generation and Editting wiht Wan series models. It offers text2image, image editting(with prompt), text2video, image2video and reference(imag...
krisyejh5.5K10
93
🪞 GPT Image 2 — Image Generation via Your ChatGPT Subscription
Generate images with GPT Image 2 (ChatGPT Images 2.0) inside Claude Code, using your existing ChatGPT Plus or Pro subscription — no separate OpenAI access, n...
Kalvin5.5K11
94
Voice Reply
Local text-to-speech using Piper voices via sherpa-onnx. 100% offline, no API keys required. Use when user asks for a voice reply, audio response, spoken answer, or wants to hear something read aloud. Supports multiple languages including German (thorsten) and English (ryan) voices. Outputs Telegram-compatible voice notes with [[audio_as_voice]] tag.
stolot0mt0m5.5K9
95
Table Image
Generate clean table images from data. Perfect for Discord/Telegram where ASCII tables look broken. Supports dark/light mode, custom styling, and auto-sizing...
Danny Shmueli5.4K17
96
公众号爆款封面生成
公众号爆款封面AI设计工具。基于全网每日收录的10w+文章数据,获取同赛道爆款封面视觉元素,通过AI分析总结高转化视觉规律,生成贴合文章内容、符合平台流量审美的封面设计方案。当用户设计公众号封面、生成封面方案、查询爆款封面规律时使用。触发词:公众号封面、封面设计、爆款封面、封面方案、封面生图。
to the moon5.4K17
97
Minimax-Multimodal-Toolkit
Use mmx to generate text, images, video, speech, and music via the MiniMax AI platform. Use when the user wants to create media content, chat with MiniMax mo...
MiniMax-AI5.1K20
98
Cheapest Image Generation
Possibly the cheapest AI image generation (~$0.0036/image). Text-to-image via the EvoLink API.
EvolinkAI5.0K6
99
VAP Media API skill for Realistic image video music
VAP Media API skill for image, video, music, and media editing through VAP. Uses VAP product keys with current Media API endpoints.
VAP Media automation flow5.0K10
100
Fal.ai API
Generate images, videos, and audio via fal.ai API (FLUX, SDXL, Whisper, etc.)
agmmnn5.0K1
Browse all 4,490 Design & Media skills

More Categories

Category listings include Claude skills with 1,000+ downloads, merged from Clawhub & SkillHub. See methodology.