デザイン・メディアのおすすめClaudeスキル6ページ目

#スキル作者ダウンロード数スター数
251Instagram Reels
AI social media video and content creation powered by CellCog. Instagram Reels, TikTok videos, Stories, carousels, social posts. Full video production from a single prompt — scripted, shot, stitched, and scored automatically.
CellCog↓ 6333★ 20
252Baoyu Image Gen
AI image generation with OpenAI GPT Image 2, Azure OpenAI, Google, OpenRouter, DashScope, Z.AI GLM-Image, MiniMax, Jimeng, Seedream, Replicate and Agnes APIs...
Jim Liu 宝玉↓ 6295★ 4
253gpt-image-2
AI图片生成技能,使用 gpt-image-2 模型根据文字描述生成高质量图片。用户安装后需要提供访问密钥才能使用。适用于:用户要求生成图片、画图、AI绘图、文生图、生成一张图等场景。
KIRIBON43567↓ 6276★ 7
254video-generator-seedance
使用火山引擎 SD1.5pro API 生成视频。支持文本到视频和图生视频,异步处理任务。
aaronzoe↓ 6222★ 2
255Video Downloader
Download and archive YouTube video content. Retrieves videos with metadata while respecting platform policies.
u_d197a013↓ 6214★ -
256Audio
Process, enhance, and convert audio files with noise removal, normalization, format conversion, transcription, and podcast workflows.
Iván↓ 6192★ 2
257Advanced QR Intelligence
Generate and read QR codes. Use when the user wants to create a QR code from text/URL, or decode/read a QR code from an image file. Supports PNG/JPG output and can read QR codes from screenshots or image files.
Omar-Khaleel↓ 6155★ 1
258Design
Executes and critiques visual design with quantified rules for hierarchy, spacing, type scale, color, and layout. Use when designing UI screens, landing pages, slides, posters, social graphics, emails, charts, or PDFs, when a design looks off, cluttered, amateur, or flat and the user cannot say why, when nothing stands out or text is hard to read, when choosing fonts, pairing typefaces, or building a palette without a design system, when adapting a screen to dark mode or mobile, or when reviewing generated HTML/CSS before delivery. Not for design tokens and component libraries (design-system) or brand strategy (branding).
Iván↓ 6154★ 6
259AI背景音乐生成
为视频、直播、播客、游戏和商业内容生成纯音乐 BGM 与背景配乐。 Use this whenever the user wants 背景音乐生成, AI背景音乐, BGM生成, 配乐生成, 纯音乐, 视频BGM, 游戏配乐, 播客配乐. Call the local aimv CLI through the Node entrypoint.
user_87b8e34f↓ 6152★ 3
260AI Song Cover Studio
Turn a song recording into a newly interpreted cover or rearranged song from a reference performance and a fresh genre, arrangement, and vocal direction. This AI song cover studio and reference-audio cover generator reinterprets a classic song as rock, acoustic, folk, jazz, Chinese-style, ballad, or a new vocal character, and reviews recognizable reference influence, arrangement freshness, vocal delivery, lyrics, pronunciation, and structure. Use it for classic-song reinterpretation, genre swaps, creator demos, tribute performances, and personal practice recordings, with one cover generation per run and honest post-result review.
beatra-ai↓ 6126★ 4
261Video Editing
Edit videos with AI background removal, color grading, upscaling, stabilization, and enhancement tools.
Iván↓ 6119★ 4
262招牌门头设计-万维易源
基于用户上传的门头照片和设计要求,进行安全审核、视觉特征分析和需求意图解析,融合生成专业的生图提示词,最终输出重绘后的门头设计效果图及后续优化建议
u_b7ff7d1b↓ 6067★ -
263Vision
Resize, crop, convert, and optimize images using ImageMagick. Use when processing photos, converting formats (PNG/WebP), compressing size, or adding watermarks.
bytesagain4↓ 5982★ 1
264Baoyu Compress Image
Compresses images to WebP (default) or PNG with automatic tool selection. Use when user asks to "compress image", "optimize image", "convert to webp", or red...
Jim Liu 宝玉↓ 5947★ -
265Fal.ai API
Generate images, videos, and audio via fal.ai API (FLUX, SDXL, Whisper, etc.)
agmmnn↓ 5946★ 1
266macOS Local Voice
Local STT and TTS on macOS using native Apple capabilities. Speech-to-text via yap (Apple Speech.framework), text-to-speech via say + ffmpeg. Fully offline, no API keys required. Includes voice quality detection and smart voice selection.
STRRL↓ 5921★ 1
267AI歌曲生成
用一句话灵感生成完整 AI 歌曲,支持中文流行歌、民谣、电子、品牌歌和创作者歌曲小样。 Use this whenever the user wants AI写歌, AI歌曲生成, 生成歌曲, 原创歌曲, 一键写歌, 中文AI写歌, AI作曲, 写一首歌. Call the local aimv CLI through the Node entrypoint.
user_87b8e34f↓ 5891★ -
268YouTube Thumbnail Maker
Create YouTube thumbnails from a video topic, title, script, key frame, portrait, product photo, or channel reference. This AI thumbnail maker compares three directions in text first, then renders the one you pick as a 16:9 image with clear visual hooks, readable hierarchy, and a headline-safe composition for explainers, reviews, tutorials, vlogs, games, podcasts, and long-form creative videos. It pairs the image with title-matching advice and refines an accepted direction into a consistent channel look. Rendering requires a Beatra account at beatra.ai.
beatra-ai↓ 5873★ 5
269Web Design
CSS implementation patterns for layout, typography, color, spacing, and responsive design. Complements ui-design (fundamentals) with code-focused examples.
wpank↓ 5868★ 3
270Minimax-Multimodal-Toolkit
Use mmx to generate text, images, video, speech, and music via the MiniMax AI platform. Use when the user wants to create media content, chat with MiniMax mo...
MiniMax-AI↓ 5845★ 20
271Music Generation
AI music generation powered by CellCog. Original instrumental and vocal tracks, 5 seconds to 10 minutes. Cinematic scores, background tracks, podcast intros, game soundtracks, ambient soundscapes, jingles, lo-fi beats, orchestral compositions, songs with lyrics. Royalty-free.
CellCog↓ 5844★ 5
272图片视频生成
可调用 banana、sora、veo 等模型生成图片视频,短剧创作者福音!可跑工作流
XuHongFeii2↓ 5775★ -
273Suno AI音乐生成
面向中文创作者的 Suno AI音乐生成工具,支持 AI写歌、歌词成曲、BGM 和参考音频改编。 Use this whenever the user wants suno, suno ai音乐, Suno AI音乐, suno ai写歌, suno中文, 中文suno, suno替代, suno音乐生成. Call the local aimv CLI through the Node entrypoint.
user_87b8e34f↓ 5747★ 5
274宠物神格鉴定-万维易源
根据用户上传的宠物图片,识别宠物特征并匹配中国传统神话神祇,生成神像风格图片及趣味TTS播报音频;适用于宠物趣味内容创作、社交分享、文创IP生成场景
u_b7ff7d1b↓ 5735★ -
275语音转文字
Download audio from any URL and transcribe it to text using Faster-Whisper. Use when asked to "transcribe this audio", "download and transcribe", "speech to text", "audio to text", "转录", "语音转文字", "音频转文字", or any task involving converting audio/video content (YouTube, Xiaoyuzhoufm, podcasts, etc.) into text. Supports YouTube, Xiaoyuzhoufm, and any site yt-dlp can extract audio from. Runs entirely locally — no cloud API needed.
user_03f2ab96↓ 5733★ 4
276Cheapest Image Generation
Possibly the cheapest AI image generation (~$0.0036/image). Text-to-image via the EvoLink API.
evolinkai↓ 5721★ 6
277人物景点打卡照生成-万维易源
上传人物照片,输入打卡地点和姿势要求,生成写实风格的景点打卡照片;适用于旅游打卡、虚拟旅游、社交媒体配图、创意合影场景
u_b7ff7d1b↓ 5715★ -
278Anthropic Frontend Design
Create distinctive, production-grade frontend interfaces that avoid generic "AI slop" aesthetics. Combines the design intelligence of UI/UX Pro Max with Anthropic's anti-slop philosophy. Use for building UI components, pages, applications, or interfaces with exceptional attention to detail and bold creative choices.
Qrucio↓ 5713★ 4
279十字绣图纸生成器-万维易源
接收任意照片生成专业的十字绣格子图纸,自动计算并标注简体中文色号、针数、用线量;适用于照片转十字绣图案、手工DIY设计、绣品材料预算场景
u_b7ff7d1b↓ 5702★ -
280image-reader
Image recognition and understanding tool. Uses a multimodal model (e.g. doubao-seed-2.0-pro, kimi-k2.5) to analyze image content and supports OCR text extrac...
simonjoe246↓ 5666★ 4
281Ppt To Video
将 PPT/PPTX 文件转换为带背景音乐的 MP4 视频。 支持智能时长(封面1.5s/正文按180字分)、翻页动画(上翻/对角线)、正文页打字机文字渐入效果。 触发词:"PPT转视频"、"PPT生成视频"、"PPT加音乐"、"将PPT做成视频"、"PPT to video"。
haoyiyong985↓ 5578★ 8
282ElevenLabs Speech-to-Text
Transcribe audio files using ElevenLabs Speech-to-Text (Scribe v2).
clawdbotborges↓ 5563★ 6
283蛋糕草图效果图生成-万维易源
面点师通过提示词和蛋糕草图生成蛋糕实际效果图。适用于烘焙设计、蛋糕定制效果图生成、草图还原、视觉丰富化场景
u_b7ff7d1b↓ 5560★ -
284Theme Factory
Curated collection of professional color and typography themes for styling artifacts — slides, docs, reports, landing pages. Use when applying visual themes to presentations, generating themed content, or creating custom brand palettes. Triggers on theme, color palette, font pairing, slide styling, presentation theme, brand colors.
wpank↓ 5541★ 3
285Unsplash
Search, browse, and download high-quality free photos from Unsplash with filtering, random selection, and detailed photo metadata access.
BrokenWatchCEO↓ 5538★ 3
286MLX STT
Speech-To-Text with MLX (Apple Silicon) and opensource models (default GLM-ASR-Nano-2512) locally.
guoqiao↓ 5530★ 2
287Image Generation
Create or revise document, PDF, web, or review images with the requested format, sharp raster output, and artifact validation.
xrow GmbH↓ 5503★ 1
288AI Video Restyler
Restyle one short video into a new visual treatment while carrying forward the source subject, action, composition, and camera intent. This AI video restyler and video style transfer workflow turns live action into anime, illustration, Chinese comic, ink, clay, paper-cut, or cyberpunk looks from one source clip and a chosen art direction, and reviews style match, subject identity, motion continuity, and source-audio result. Use it for live action to anime, brand visual refresh, Chinese-comic looks, fashion films, music visuals, and creator experiments, with one dominant visual change per run and honest post-result review.
beatra-ai↓ 5501★ 6
289视频内容智能分析工具
分析视频内容,支持今日头条、抖音、B站、西瓜视频等平台。自动下载视频、语音转文字(faster-whisper)、繁简转换、生成内容摘要。当用户提供视频链接并要求分析视频内容、总结视频、看看视频讲什么时使用此技能。
user_09d1a78f↓ 5492★ 19
290Transcribe
Transcribe audio files to text using local Whisper (Docker). Use when receiving voice messages, audio files (.mp3, .m4a, .ogg, .wav, .webm), or when asked to transcribe audio content.
javicasper↓ 5491★ 2
291Excalidraw Diagram Generator
Generate Excalidraw diagrams from natural language descriptions. Use when asked to "create a diagram", "make a flowchart", "visualize a process", "draw a sys...
Morpheous↓ 5478★ 7
292Tts
Convert text to speech using Hume AI (or OpenAI) API. Use when the user asks for an audio message, a voice reply, or to hear something "of vive voix".
AMSTKO↓ 5475★ 1
293Canva Connect
Manage Canva designs, assets, and folders via the Connect API. WHAT IT CAN DO: - List/search/organize designs and folders - Export finished designs (PNG/PDF/JPG) - Upload images to asset library - Autofill brand templates with data - Create blank designs (doc/presentation/whiteboard/custom) WHAT IT CANNOT DO: - Add content to designs (text, shapes, elements) - Edit existing design content - Upload documents (images only) - AI design generation Best for: asset pipelines, export automation, organization, template autofill. Triggers: /canva, "upload to canva", "export design", "list my designs", "canva folder".
coolmanns↓ 5466★ 4
294AI歌曲
AI歌曲创作工具,支持一句话写歌、歌词成曲、中文流行歌、品牌歌和歌曲小样生成。 Use this whenever the user wants AI歌曲, ai歌曲, AI歌曲生成, AI唱歌, AI写歌, AI作曲, 生成歌曲, 写一首歌. Call the local aimv CLI through the Node entrypoint.
user_87b8e34f↓ 5466★ 1
295文本识别OCR
兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。
xby-skill↓ 5466★ 1
296AudioPod
Use AudioPod AI's API for audio processing tasks including AI music generation (text-to-music, text-to-rap, instrumentals, samples, vocals), stem separation, text-to-speech, noise reduction, speech-to-text transcription, speaker separation, and media extraction. Use when the user needs to generate music/songs/rap from text, split a song into stems/vocals/instruments, generate speech from text, clean up noisy audio, transcribe audio/video, or extract audio from YouTube/URLs. Requires AUDIOPOD_API_KEY env var or pass api_key directly.
Rakesh Roushan↓ 5458★ 5
297Video Analyzer
视频分析处理 — 本地视频反编译分析工具。将视频拆解为时间轴剧本、语音转文字、场景分析、跨模态关联和精华摘要,支持多ASR引擎切换(Whisper/Paraformer/SenseVoice)、中文NLP增强、PaddleOCR中文识别。v4.0 新增短视频平台适配(抖音/快手/B站/视频号)和自动剪辑建议(高光检测/冗余标记/EDL导出/字幕样式)。v4.1 新增tiny模型优先体验(75MB低门槛)、说话人分离质量评分、剪映draft.json导出。v4.2 新增场景管理(detect→slice一条链)、短视频爆款预测、实时直播分析(流式ASR+敏感词检测)。v4.3 新增纯音频输入(mp3/m4a/wav播客与录音)、批量队列(SQLite+硬件档位并发)、GPU自动加速(CT2 int8量化)、ASR配置统一(--asr-engine单参数)。
fyniujin↓ 5434★ 2
298即梦-图生视频 CLI 技能,让你的智能体通过 即梦 CLI 生成视频
Use when the user provides image, video, or audio references and wants Dreamina 即梦 video generation through `image2video`, `frames2video`, `multiframe2video`, or `multimodal2video`. Covers CLI v1.4.14 mode routing, required video resolution, Seedance model sets, input limits, and async terminal statuses.
user_c1d16043↓ 5414★ 4
299宠物写真-万维易源
基于用户上传的宠物原图,通过视觉特征提取、提示词智能构建与图生图渲染,生成高保真保留宠物本体特征的拟人化高清写真图片;适用于宠物IP形象设计、趣味萌宠拟人化、宠物周边定制场景
u_38e1d485↓ 5396★ 1
300Voice
Convert text to speech using Microsoft Edge's TTS engine with customizable voices, direct playback, and automatic temporary file cleanup.
zhaov↓ 5381★ -
デザイン・メディアのスキル13,243件をすべて見る →

その他のカテゴリ

カテゴリ一覧にはダウンロード数5,000以上のClaudeスキルを掲載し、データはClawhubとSkillHubを統合しています。詳しくはデータの方法論をご覧ください。