Best Design & Media Claude SkillsPage 3

#SkillAuthorDownloadsStars
101
Vision
Resize, crop, convert, and optimize images using ImageMagick. Use when processing photos, converting formats (PNG/WebP), compressing size, or adding watermarks.
bytesagain44.9K1
102
ElevenLabs Speech-to-Text
Transcribe audio files using ElevenLabs Speech-to-Text (Scribe v2).
clawdbotborges4.8K6
103
Brand Cog
AI brand identity design powered by CellCog. Brand kits, color palettes, typography, brand guidelines, logo design, visual identity, social media assets, web...
CellCog4.8K4
104
Seedance 2.0 prompt-engineering skill
Generate precise, timecoded Seedance 2.0 prompts integrating multimodal inputs with asset mapping for controlled 4-15s video creation and editing.
Daniil Burykin4.8K27
105
MLX STT
Speech-To-Text with MLX (Apple Silicon) and opensource models (default GLM-ASR-Nano-2512) locally.
guoqiao4.7K2
106
AudioPod
Use AudioPod AI's API for audio processing tasks including AI music generation (text-to-music, text-to-rap, instrumentals, samples, vocals), stem separation, text-to-speech, noise reduction, speech-to-text transcription, speaker separation, and media extraction. Use when the user needs to generate music/songs/rap from text, split a song into stems/vocals/instruments, generate speech from text, clean up noisy audio, transcribe audio/video, or extract audio from YouTube/URLs. Requires AUDIOPOD_API_KEY env var or pass api_key directly.
Rakesh10024.6K5
107
AI Image & Video Generator — GPT Image 2, Seedance, ComfyUI
Generate images from text with multi-provider routing — supports GPT Image 2, Nanobanana 2, Midjourney V8.1 & Seedance 2
jau1234.6K8
108
小红书内容创作
Generate Xiaohongshu (小红书/RED) content optimized for the platform's CES algorithm. Use when: (1) creating xiaohongshu/小红书 posts, (2) writing Chinese social m...
3727684984.6K1
109
Tts
Convert text to speech using Hume AI (or OpenAI) API. Use when the user asks for an audio message, a voice reply, or to hear something "of vive voix".
AMSTKO4.6K1
110
Volcengine Ai Image Generation
Image generation workflow on Volcengine AI services. Use when users need text-to-image, style variants, prompt refinement, or deterministic image generation parameters and troubleshooting.
cinience4.6K3
111
Ui Ux Pro Max
UI/UX design intelligence. 50 styles, 21 palettes, 50 font pairings, 20 charts, 9 stacks (React, Next.js, Vue, Svelte, SwiftUI, React Native, Flutter, Tailwind, shadcn/ui). Actions: plan, build, create, design, implement, review, fix, improve, optimize, enhance, refactor, check UI/UX code. Projects: website, landing page, dashboard, admin panel, e-commerce, SaaS, portfolio, blog, mobile app, .html, .tsx, .vue, .svelte. Elements: button, modal, navbar, sidebar, card, table, form, chart. Styles: glassmorphism, claymorphism, minimalism, brutalism, neumorphism, bento grid, dark mode, responsive, skeuomorphism, flat design. Topics: color palette, accessibility, animation, layout, typography, font pairing, spacing, hover, shadow, gradient. Integrations: shadcn/ui MCP for component search and examples.
ahump204.5K13
112
Transcribe audio files via OpenRouter using audio-capable models
Transcribe audio files via OpenRouter using audio-capable models (Gemini, GPT-4o-audio, etc).
obviyus4.5K4
113
Music Cog
AI music generation powered by CellCog. Original instrumental and vocal tracks, 5 seconds to 10 minutes. Cinematic scores, background tracks, podcast intros,...
CellCog4.4K4
114
去水印
抖音、小红书、豆包、千问等平台的图片/视频去水印解析。当用户需要去除视频/图片水印、获取无水印直链、调用解析API、查询价格时使用此skill。
user_bfa2fd914.4K15
115
Designer
Create logos, interfaces, and visual systems with principles of hierarchy, branding, and usability.
Iván4.4K4
116
Cinematic Script Writer
Create professional cinematic scripts for AI video generation with character consistency and cinematography knowledge. Use when the user wants to write a cinematic script, create story contexts with characters, generate image prompts for AI video tools (Midjourney, Sora, Veo), or needs cinematography guidance (camera angles, lighting, color grading). Also use for character consistency sheets, voice profiles, anachronism detection, and saving scripts to Google Drive.
praveenspeaks4.3K9
117
Video Editor
Perform video editing tasks with ffmpeg, including cutting, merging, converting formats, extracting audio, adding subtitles, resizing, cropping, adjusting sp...
cntuang4.3K2
118
Voice
Convert text to speech using Microsoft Edge's TTS engine with customizable voices, direct playback, and automatic temporary file cleanup.
zhaov4.3K-
119
captions
Use when captions, subtitles, or the spoken text of a YouTube video is needed — even if not explicitly requested: pasted video links or IDs, requests to read...
Rohit Das4.3K1
120
LibTV API Skills
agent-im 会话技能 - 通过 liblib.tv 的 AI 能力生成和编辑图片/视频。覆盖场景包括:生成(文生图、文生视频、图生视频、做动画、画一个xxx、来段xxx)、编辑修改(把xxx换成yyy、去掉xxx、加上xxx、改成xxx、调整xxx、局部修改、改镜头)、风格转换(风格迁移、转绘、换风格)、视...
Frank (Haofan) Wang4.3K10
121
UI Design
Comprehensive UI design skill covering fundamentals, patterns, and anti-patterns. Layout, typography, color, spacing, accessibility, motion, and component design. Use when building any web interface, reviewing design quality, or creating distinctive UIs.
wpank4.3K-
122
Douyin Messager | 抖音私信助手
Douyin DMs and video/note comment assistant. 抖音私信与视频/图文评论助手;可读私信、分析评论区,评论/回复等写入前必须确认。
Morois4.3K14
123
Baoyu Image Gen
AI image generation with OpenAI GPT Image 2, Azure OpenAI, Google, OpenRouter, DashScope, Z.AI GLM-Image, MiniMax, Jimeng, Seedream, Replicate and Agnes APIs...
Jim Liu 宝玉4.3K1
124
Mlx Whisper
Local speech-to-text with MLX Whisper (Apple Silicon optimized, no API key).
Kevin37Li4.2K1
125
抖音视频转文字
抖音视频转文字。用户发抖音链接或视频文件,自动转录为带标点的中文文本。 触发词:抖音、转文字、转录、视频转文本、douyin、transcribe
junweiren98-rgb4.2K2
126
Baoyu Comic
Knowledge comic creator supporting multiple art styles and tones. Creates original educational comics with detailed panel layouts and batch-capable image gen...
Jim Liu 宝玉4.2K3
127
subtitles
Use when subtitles or the spoken text of a YouTube video is needed: pasted video links or IDs, requests to translate a video, read along, follow foreign-lang...
Rohit Das4.2K-
128
TG Voice Whisper Transcriber
Automation skill for TG Voice Whisper Transcriber.
RigdenDjapo4.2K1
129
Agent Selfie
AI agent self-portrait generator. Create avatars, profile pictures, and visual identity using Gemini image generation. Supports mood-based generation, season...
김덕환4.2K8
130
OpenSpec
Spec-driven development with OpenSpec CLI. Use when building features, migrations, refactors, or any structured development work. Manages proposal → specs → design → tasks → implementation workflows. Supports custom schemas (TDD, rapid, etc.). Trigger on requests involving feature planning, spec writing, change management, or when /opsx commands are mentioned.
jcorrego4.2K-
131
Ad Context Protocol (AdCP) Advertising
Automate advertising campaigns with AI. Create ads, buy media, manage ad budgets, discover ad inventory, run display ads, video ads, CTV campaigns, and optimize ad performance. Perfect for marketing automation, programmatic advertising, media buying, ad management, campaign optimization, creative management, and performance tracking. Launch Facebook ads, Google ads, display advertising, video marketing, and multi-channel campaigns using natural language. Supports ad targeting, audience segmentation, ROI tracking, and automated bidding.
edyyy624.2K8
132
Xiaohongshu Content Creator
小紅書內容創作技能 - 爆款筆記標題、內容模板、hashtag策略、發布時間優化。適合博主、品牌、MCN。觸發詞:小紅書、紅書、xiaohongshu、筆記、種草、爆款、小紅書文案、筆記標題。
isaacloi1995-dot4.1K4
133
Vocal Chat
Handles voice-to-voice conversations on WhatsApp. Automatically transcribes incoming audio and responds with local TTS audio. Use when the user wants to "talk" instead of type.
Rubén Fernández Boullón4.1K12
134
PRD to Prototype
从产品需求到可交互原型的完整工作流。当用户表达产品想法(如'我想做一个...'、'帮我设计...')时触发。支持:1) 零提问直出PRD 2) 平台选择确认 3) 生成Awwwards级别的高保真HTML/Tailwind原型(移动端/PC端)。端到端产品设计流程。
zengxming454.1K13
135
Anthropic Frontend Design
Create distinctive, production-grade frontend interfaces that avoid generic "AI slop" aesthetics. Combines the design intelligence of UI/UX Pro Max with Anthropic's anti-slop philosophy. Use for building UI components, pages, applications, or interfaces with exceptional attention to detail and bold creative choices.
Qrucio4.1K2
136
HeyGen AI Avatar Video (Lite)
Create AI digital human videos with HeyGen API. Free starter guide.
Ju Chun Ko4.0K7
137
Transcribe
Transcribe audio files to text using local Whisper (Docker). Use when receiving voice messages, audio files (.mp3, .m4a, .ogg, .wav, .webm), or when asked to transcribe audio content.
javicasper4.0K2
138
Sherpa ONNX TTS
Local text-to-speech via sherpa-onnx (offline, no cloud)
Daniel Sinewe4.0K1
139
Music Generation
Generate AI music with optimized prompts, style control, and production-ready audio output.
Iván4.0K5
140
TTS WhatsApp
Send high-quality text-to-speech voice messages on WhatsApp in 40+ languages with automatic delivery
hopyky4.0K9
141
Gemini Yt Video Transcript
Create a verbatim transcript for a YouTube URL using Google Gemini (speaker labels, paragraph breaks; no time codes). Use when the user asks to transcribe a YouTube video or wants a clean transcript (no timestamps).
Oliver Drobnik3.9K3
142
Swiftui Liquid Glass
Implement, review, or improve SwiftUI features using the iOS 26+ Liquid Glass API. Use when asked to adopt Liquid Glass in new SwiftUI UI, refactor an existing feature to Liquid Glass, or review Liquid Glass usage for correctness, performance, and design alignment.
Peter Steinberger3.9K16
143
AI图片生成
AI生成图片,支持Kolors/FLUX/Qwen-Image等模型(需SiliconFlow API)
asdas3.9K-
144
Gemini STT
Transcribe audio files using Google's Gemini API or Vertex AI
araa473.9K2
145
ElevenLabs Music
Generate music from text prompts using ElevenLabs Eleven Music API. Use when creating songs, soundtracks, jingles, lullabies, or any audio music from descriptions. Supports vocals with AI-generated lyrics, instrumental tracks, and multiple genres/styles. Requires paid ElevenLabs plan.
clawdbotborges3.9K1
146
video-transcript
Use when video content needs to be extracted as text: pasted YouTube links or IDs, requests to transcribe, summarize, quote, translate, convert video to text...
Rohit Das3.9K3
147
Design
Auto-learns your visual preferences. Adapts to UI, graphics, video, and any creative work.
Iván3.8K4
148
Wonda
Using the Wonda CLI to generate images, videos, music, and audio from the terminal — plus LinkedIn, Reddit, and X/Twitter research and automation
Degaus AI3.8K1
149
Speech To Text
Transcribe audio to text with Whisper models via inference.sh CLI. Models: Fast Whisper Large V3, Whisper V3 Large. Capabilities: transcription, translation,...
Ömer Karışman3.8K-
150
MoltTok
After installing, tell your agent: Activate MoltTok. Creative expression platform where AI agents post ASCII art, SVG, HTML, p5.js, and poetry to a live feed.
tristankaiburrell-code3.7K2
Browse all 4,490 Design & Media skills

More Categories

Category listings include Claude skills with 1,000+ downloads, merged from Clawhub & SkillHub. See methodology.