Best Design & Media Claude SkillsPage 10

#SkillAuthorDownloadsStars
451
Figma
Read files, manage comments, extract design tokens, download images, and create webhooks in Figma via the Figma REST API. Use this skill when users want to i...
Jay1.9K32
452
ElevenLabs
Create and manage voices, speech synthesis, audio projects, and conversational AI agents in ElevenLabs via the ElevenLabs API. Use this skill when users want...
Jay1.9K31
453
AssemblyAI Transcriber
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
xenofex71.9K-
454
SVG
Create and optimize SVG graphics with proper viewBox, accessibility, and CSS styling.
Iván1.9K2
455
mediaproc
Process media files (video, audio, images) via a locked-down SSH container with ffmpeg, sox, and imagemagick. Use when the user wants to transcode video, pro...
Ciprian Mandache1.9K-
456
Screenshot Skill
Capture screenshots on Windows using mss and Pillow. Provides full-screen, region, and multi-monitor capture with output as PIL Image, PNG file, or base64 st...
sunrdd1.9K-
457
Alicloud Ai Image Qwen Image
Generate images with Model Studio DashScope SDK using Qwen Image generation models (qwen-image, qwen-image-plus, qwen-image-max and snapshots). Use when impl...
cinience1.9K2
458
音乐生成
Generate custom music tracks (vocal or instrumental) via OhYesAI asynchronously.
bajie-git1.9K3
459
Ecommerce Image Asset Generator
Plan and generate ecommerce image assets that actually support conversion. Use when teams need to decide which product, PDP, marketplace, promo, or ad images...
LeroyCreates1.9K1
460
z-card-image
生成配图、封面图、卡片图、文字海报、公众号文章封面图、微信公众号头图、X 风格帖子分享图、帖子长图、社媒帖子长图。适用于帖子类型数据、post data、social posts、tweet/thread、转发推文、转发帖子、小绿书配图、图片封面、card image。
Jinx1.9K3
461
Nano Banana Pro (Morfeo)
Generate and edit images using Google's Nano Banana Pro (Gemini 3 Pro Image) API. Use when the user asks to generate, create, edit, modify, change, alter, or update images. Also use when user references an existing image file and asks to modify it in any way (e.g., "modify this image", "change the background", "replace X with Y"). Supports both text-to-image generation and image-to-image editing with configurable resolution (1K default, 2K, or 4K for high resolution). DO NOT read the image file first - use this skill directly with the --input-image parameter.
Paul de Lavallaz1.9K3
462
Clawpix
AI image sharing platform where agents post and discover AI-generated art. Register, authenticate, and share your creations with the world.
ryan3211.9K1
463
Dlazy Video Generate
Video generation skill. Automatically selects the best dlazy CLI video model based on the prompt.
dlazy1.9K-
464
混元生视频能力
腾讯混元生视频API - 支持文生视频、图生视频、视频风格化
Tony1.9K1
465
MiniMax Speech 2.8
Manage MiniMax Speech 2.8 TTS requests, voice catalog lookups, and precise voice/audio configuration using MiniMax API via CLI or script.
wingchiu1.8K-
466
IMA AI Video Generator — Short & Promo Video, Text to Video, Image to Video Generation
AI video generator with premier models: Wan 2.6, Kling O1/2.6, Google Veo 3.1, Sora 2 Pro, Pixverse V5.5, Hailuo 2.0/2.3, SeeDance 1.5 Pro, Vidu Q2. Video ge...
Dai Shuo1.8K2
467
Architecture Diagram
AI architecture diagram generator supporting Mermaid charts. Generate system architecture, cloud architecture, neural network, graph theory, flowchart, ER di...
SMS1.8K6
468
Upload audio to AIOZ Stream
Quick upload audio to AIOZ Stream API. Create audio objects with default or custom encoding configurations, upload the file, complete the upload, then return the audio link to the user.
vinhbui30041.8K-
469
Dlazy Generate
A comprehensive generation skill. Can generate images, videos, and audio by automatically selecting the appropriate dlazy CLI model.
dlazy1.8K1
470
mac-node-snapshot
A robust, permission-friendly method to capture macOS screens via OpenClaw screen.record. Ideal for headless environments or ensuring capture reliability.
taozhe61.8K-
471
Grok Image Cli
Generate and edit images via Grok API from the command line. Cross-platform secure credential storage for xAI API key. Supports batch generation, aspect rati...
cyberash-dev1.8K-
472
MLX TTS
Text-To-Speech with MLX (Apple Silicon) and opensource models (default QWen3-TTS) locally.
guoqiao1.8K-
473
🎤 Transcribe audio files using Qwen ASR. 千问STT
Transcribe audio files using Qwen ASR (千问STT). Use when the user sends voice messages and wants them converted to text.
Alone1.8K2
474
xhs Agent
xhs 全流程助手,覆盖小红书内容策划、文案与标题生成、封面制作、笔记发布及日常运营管理。适用于写笔记、生成标题/封面、发布或保存草稿、站内搜索、评论互动(点赞/收藏/回复)等小红书相关任务。支持从内容创作到发布执行的一站式流程;封面 AI 生图可选配置 GEMINI_API_KEY、IMG_API_KEY 或...
Lucas1.8K5
475
TitleClash
Compete in TitleClash - write creative titles for images and win votes. Use when user wants to play TitleClash, submit titles, or check competition results.
appback1.8K2
476
Article Illustrator
When the user wants to add illustrations to an article or blog post. Triggers on: "illustrate article", "add images to article", "generate illustrations", "article images", or requests to visually enhance written content. Analyzes article structure, identifies positions for visual aids, and generates illustrations using a Type x Style two-dimension approach.
wpank1.8K2
477
Pencil Renderer
Render DNA codes to Pencil .pen frames. Does ONE thing well. Input: DNA code + component type (hero, card, form, etc.) Output: .pen frame ID + screenshot Use when: design-exploration or other orchestrators need to render visual proposals using Pencil MCP backend.
Jcwen1.8K-
478
Cinematic Script Writer
Create professional cinematic scripts for AI video generation with character consistency and cinematography knowledge. Use when the user wants to write a cinematic script, create story contexts with characters, generate image prompts for AI video tools (Midjourney, Sora, Veo), or needs cinematography guidance (camera angles, lighting, color grading). Also use for character consistency sheets, voice profiles, anachronism detection, and saving scripts to Google Drive.
praveenspeaks1.8K2
479
Distinctive Design Systems
Patterns for creating design systems with personality and distinctive aesthetics. Covers aesthetic documentation, color token architecture, typography systems, layered surfaces, and motion. Use when building design systems that go beyond generic templates. Triggers on design system, design tokens, aesthetic, color palette, typography, CSS variables, tailwind config.
wpank1.8K-
480
Baoyu Danger Gemini Web
Generates images and text via reverse-engineered Gemini Web API. Supports text generation, image generation from prompts, reference images for vision input,...
Jim Liu 宝玉1.8K1
481
URL to PNG
Convert URL to PNG suitable for mobile reading.
guoqiao1.8K1
482
Diagram Generator
Use this skill any time the user wants to create diagrams, flowcharts, or visual structures. This includes: architecture diagrams, mind maps, org charts, use...
AnyGenIO1.8K1
483
Gstack Openclaw
世界顶级思维合集 —— 融合Google Staff Engineer、Martin Fowler/Kent Beck/Jeff Dean工程思维、Paul Graham/Sam Altman创业思维、Elon Musk创新思维、Stripe/Airbnb设计思维。v2.5.10:移除install.sh以完全消...
leo-jiqimao1.8K1
484
aliyun-image
阿里云百炼图像生成、编辑与翻译。文生图:根据文本生成图像,支持复杂文字渲染。图像编辑:单图编辑、多图融合、风格迁移、物体增删。图像翻译:翻译图像中的文字,保留原始排版,支持11种源语言和14种目标语言。触发词:生成图片、AI作画、文生图、图像编辑、修图、换背景、风格迁移、多图融合、图像翻译、图片翻译。模型:qwen-image-plus(默认)、qwen-image-max、qwen-image-edit-plus(默认)、qwen-image-edit-max、qwen-mt-image。
StanleyChanH1.8K1
485
AI Video Editor
Raw assets + Requirements in, result video out. Edit any type of video with Sparki (the video agent powered by Gemini multimodal AI).
BoShen1.8K4
486
Faster Whisper Transcription
Transcribes local voice messages to text using Faster Whisper models for fast, privacy-focused speech recognition on audio files.
Khalid Almuraee1.8K-
487
Vibe Design
Create visual designs with AI tools. Covers prompting for UI/graphics, Midjourney techniques, Figma AI workflow, and iteration patterns.
Iván1.8K10
488
Video analyze by doubao2.0
使用豆包2.0模型解析视频。当需要执行分析视频内容等需要理解视频视觉信息时调用该技能。你必须在持有本地视频路径或网络视频链接时才能调用该技能
danielcy1.8K2
489
Grok Imagine Image Pro
Generates and edits high-quality PNG images via xAI Grok/Flux API using prompts, styles, aspect ratios, and batch processing with base64 output.
NixeiFoit1.8K3
490
Logo Creator
Engineer professional-grade brand logos using geometric primitives and negative space — generates minimalist, scalable vector-style marks via muapi.ai
Anil Chandra Naidu Matcha1.8K1
491
Speechall command-line tool for fast speech-to-text transcription using multiple providers
Install and use the speechall CLI tool for speech-to-text transcription. Use when the user wants to: (1) transcribe audio or video files to text, (2) install speechall on macOS or Linux, (3) list available STT models and their capabilities, (4) use speaker diarization, subtitles, or other transcription features from the terminal. Triggers on mentions of speechall, audio transcription CLI, or speech-to-text from the command line.
atacan1.8K-
492
Publish Skill Final
Design beautiful interfaces using 16+ design systems including Material You, Fluent Design, Apple HIG, Ant Design, Carbon Design, Shopify Polaris, Minimalism...
azzar budiyanto1.8K4
493
Google Veo
Generate videos with Google Veo models via inference.sh CLI. Models: Veo 3.1, Veo 3.1 Fast, Veo 3, Veo 3 Fast, Veo 2. Capabilities: text-to-video, cinematic...
Ömer Karışman1.8K2
494
Dlazy Jimeng I2v First
Generate dynamic videos based on a single first frame image and prompts using Jimeng.
dlazy1.8K-
495
Draw.io Skill
Create and edit draw.io diagrams through the configured drawio MCP server, including flowcharts, architecture diagrams, ML model diagrams, Chinese labels, an...
yyq20241.8K-
496
Claw Fm
Submit and manage music on claw.fm - the AI radio station. Use when submitting tracks, checking artist stats, engaging with comments, or managing your claw.fm presence. Triggers on "claw.fm", "submit track", "AI radio", "music submission", or artist profile management.
rawgroundbeef1.8K2
497
Photos
Organize, index, and search local photo libraries with AI-powered metadata and safe file handling.
Iván1.8K2
498
seedance-creator
This skill should be used when the user asks to "generate video prompts", "create Seedance prompts", "write video descriptions", mentions "Seedance", "seedan...
chenke1.8K1
499
Dlazy Doubao Tts
Synthesize text into natural and fluent speech using Doubao TTS.
dlazy1.8K-
500
Alicloud Ai Audio Tts
Generate human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash, qwen3-tts-instruct-flash). Use when converting text to speech,...
cinience1.8K-
Browse all 4,490 Design & Media skills

More Categories

Category listings include Claude skills with 1,000+ downloads, merged from Clawhub & SkillHub. See methodology.