| # | Skill | Author | Downloads | Stars |
|---|---|---|---|---|
| 251 | Instagram Reels AI social media video and content creation powered by CellCog. Instagram Reels, TikTok videos, Stories, carousels, social posts. Full video production from a single prompt — scripted, shot, stitched, and scored automatically. | CellCog | ↓ 6.3K | ★ 20 |
| 252 | Baoyu Image Gen AI image generation with OpenAI GPT Image 2, Azure OpenAI, Google, OpenRouter, DashScope, Z.AI GLM-Image, MiniMax, Jimeng, Seedream, Replicate and Agnes APIs... | Jim Liu 宝玉 | ↓ 6.3K | ★ 4 |
| 253 | gpt-image-2 AI图片生成技能,使用 gpt-image-2 模型根据文字描述生成高质量图片。用户安装后需要提供访问密钥才能使用。适用于:用户要求生成图片、画图、AI绘图、文生图、生成一张图等场景。 | KIRIBON43567 | ↓ 6.3K | ★ 7 |
| 254 | video-generator-seedance 使用火山引擎 SD1.5pro API 生成视频。支持文本到视频和图生视频,异步处理任务。 | aaronzoe | ↓ 6.2K | ★ 2 |
| 255 | Video Downloader Download and archive YouTube video content. Retrieves videos with metadata while respecting platform policies. | u_d197a013 | ↓ 6.2K | ★ - |
| 256 | Audio Process, enhance, and convert audio files with noise removal, normalization, format conversion, transcription, and podcast workflows. | Iván | ↓ 6.2K | ★ 2 |
| 257 | Advanced QR Intelligence Generate and read QR codes. Use when the user wants to create a QR code from text/URL, or decode/read a QR code from an image file. Supports PNG/JPG output and can read QR codes from screenshots or image files. | Omar-Khaleel | ↓ 6.2K | ★ 1 |
| 258 | Design Executes and critiques visual design with quantified rules for hierarchy, spacing, type scale, color, and layout. Use when designing UI screens, landing pages, slides, posters, social graphics, emails, charts, or PDFs, when a design looks off, cluttered, amateur, or flat and the user cannot say why, when nothing stands out or text is hard to read, when choosing fonts, pairing typefaces, or building a palette without a design system, when adapting a screen to dark mode or mobile, or when reviewing generated HTML/CSS before delivery. Not for design tokens and component libraries (design-system) or brand strategy (branding). | Iván | ↓ 6.2K | ★ 6 |
| 259 | AI背景音乐生成 为视频、直播、播客、游戏和商业内容生成纯音乐 BGM 与背景配乐。 Use this whenever the user wants 背景音乐生成, AI背景音乐, BGM生成, 配乐生成, 纯音乐, 视频BGM, 游戏配乐, 播客配乐. Call the local aimv CLI through the Node entrypoint. | user_87b8e34f | ↓ 6.2K | ★ 3 |
| 260 | AI Song Cover Studio Turn a song recording into a newly interpreted cover or rearranged song from a reference performance and a fresh genre, arrangement, and vocal direction. This AI song cover studio and reference-audio cover generator reinterprets a classic song as rock, acoustic, folk, jazz, Chinese-style, ballad, or a new vocal character, and reviews recognizable reference influence, arrangement freshness, vocal delivery, lyrics, pronunciation, and structure. Use it for classic-song reinterpretation, genre swaps, creator demos, tribute performances, and personal practice recordings, with one cover generation per run and honest post-result review. | beatra-ai | ↓ 6.1K | ★ 4 |
| 261 | Video Editing Edit videos with AI background removal, color grading, upscaling, stabilization, and enhancement tools. | Iván | ↓ 6.1K | ★ 4 |
| 262 | 招牌门头设计-万维易源 基于用户上传的门头照片和设计要求,进行安全审核、视觉特征分析和需求意图解析,融合生成专业的生图提示词,最终输出重绘后的门头设计效果图及后续优化建议 | u_b7ff7d1b | ↓ 6.1K | ★ - |
| 263 | Vision Resize, crop, convert, and optimize images using ImageMagick. Use when processing photos, converting formats (PNG/WebP), compressing size, or adding watermarks. | bytesagain4 | ↓ 6.0K | ★ 1 |
| 264 | Baoyu Compress Image Compresses images to WebP (default) or PNG with automatic tool selection. Use when user asks to "compress image", "optimize image", "convert to webp", or red... | Jim Liu 宝玉 | ↓ 5.9K | ★ - |
| 265 | Fal.ai API Generate images, videos, and audio via fal.ai API (FLUX, SDXL, Whisper, etc.) | agmmnn | ↓ 5.9K | ★ 1 |
| 266 | macOS Local Voice Local STT and TTS on macOS using native Apple capabilities. Speech-to-text via yap (Apple Speech.framework), text-to-speech via say + ffmpeg. Fully offline, no API keys required. Includes voice quality detection and smart voice selection. | STRRL | ↓ 5.9K | ★ 1 |
| 267 | AI歌曲生成 用一句话灵感生成完整 AI 歌曲,支持中文流行歌、民谣、电子、品牌歌和创作者歌曲小样。 Use this whenever the user wants AI写歌, AI歌曲生成, 生成歌曲, 原创歌曲, 一键写歌, 中文AI写歌, AI作曲, 写一首歌. Call the local aimv CLI through the Node entrypoint. | user_87b8e34f | ↓ 5.9K | ★ - |
| 268 | YouTube Thumbnail Maker Create YouTube thumbnails from a video topic, title, script, key frame, portrait, product photo, or channel reference. This AI thumbnail maker compares three directions in text first, then renders the one you pick as a 16:9 image with clear visual hooks, readable hierarchy, and a headline-safe composition for explainers, reviews, tutorials, vlogs, games, podcasts, and long-form creative videos. It pairs the image with title-matching advice and refines an accepted direction into a consistent channel look. Rendering requires a Beatra account at beatra.ai. | beatra-ai | ↓ 5.9K | ★ 5 |
| 269 | Web Design CSS implementation patterns for layout, typography, color, spacing, and responsive design. Complements ui-design (fundamentals) with code-focused examples. | wpank | ↓ 5.9K | ★ 3 |
| 270 | Minimax-Multimodal-Toolkit Use mmx to generate text, images, video, speech, and music via the MiniMax AI platform. Use when the user wants to create media content, chat with MiniMax mo... | MiniMax-AI | ↓ 5.8K | ★ 20 |
| 271 | Music Generation AI music generation powered by CellCog. Original instrumental and vocal tracks, 5 seconds to 10 minutes. Cinematic scores, background tracks, podcast intros, game soundtracks, ambient soundscapes, jingles, lo-fi beats, orchestral compositions, songs with lyrics. Royalty-free. | CellCog | ↓ 5.8K | ★ 5 |
| 272 | 图片视频生成 可调用 banana、sora、veo 等模型生成图片视频,短剧创作者福音!可跑工作流 | XuHongFeii2 | ↓ 5.8K | ★ - |
| 273 | Suno AI音乐生成 面向中文创作者的 Suno AI音乐生成工具,支持 AI写歌、歌词成曲、BGM 和参考音频改编。 Use this whenever the user wants suno, suno ai音乐, Suno AI音乐, suno ai写歌, suno中文, 中文suno, suno替代, suno音乐生成. Call the local aimv CLI through the Node entrypoint. | user_87b8e34f | ↓ 5.7K | ★ 5 |
| 274 | 宠物神格鉴定-万维易源 根据用户上传的宠物图片,识别宠物特征并匹配中国传统神话神祇,生成神像风格图片及趣味TTS播报音频;适用于宠物趣味内容创作、社交分享、文创IP生成场景 | u_b7ff7d1b | ↓ 5.7K | ★ - |
| 275 | 语音转文字 Download audio from any URL and transcribe it to text using Faster-Whisper. Use when asked to "transcribe this audio", "download and transcribe", "speech to text", "audio to text", "转录", "语音转文字", "音频转文字", or any task involving converting audio/video content (YouTube, Xiaoyuzhoufm, podcasts, etc.) into text. Supports YouTube, Xiaoyuzhoufm, and any site yt-dlp can extract audio from. Runs entirely locally — no cloud API needed. | user_03f2ab96 | ↓ 5.7K | ★ 4 |
| 276 | Cheapest Image Generation Possibly the cheapest AI image generation (~$0.0036/image). Text-to-image via the EvoLink API. | evolinkai | ↓ 5.7K | ★ 6 |
| 277 | 人物景点打卡照生成-万维易源 上传人物照片,输入打卡地点和姿势要求,生成写实风格的景点打卡照片;适用于旅游打卡、虚拟旅游、社交媒体配图、创意合影场景 | u_b7ff7d1b | ↓ 5.7K | ★ - |
| 278 | Anthropic Frontend Design Create distinctive, production-grade frontend interfaces that avoid generic "AI slop" aesthetics. Combines the design intelligence of UI/UX Pro Max with Anthropic's anti-slop philosophy. Use for building UI components, pages, applications, or interfaces with exceptional attention to detail and bold creative choices. | Qrucio | ↓ 5.7K | ★ 4 |
| 279 | 十字绣图纸生成器-万维易源 接收任意照片生成专业的十字绣格子图纸,自动计算并标注简体中文色号、针数、用线量;适用于照片转十字绣图案、手工DIY设计、绣品材料预算场景 | u_b7ff7d1b | ↓ 5.7K | ★ - |
| 280 | image-reader Image recognition and understanding tool. Uses a multimodal model (e.g. doubao-seed-2.0-pro, kimi-k2.5) to analyze image content and supports OCR text extrac... | simonjoe246 | ↓ 5.7K | ★ 4 |
| 281 | Ppt To Video 将 PPT/PPTX 文件转换为带背景音乐的 MP4 视频。 支持智能时长(封面1.5s/正文按180字分)、翻页动画(上翻/对角线)、正文页打字机文字渐入效果。 触发词:"PPT转视频"、"PPT生成视频"、"PPT加音乐"、"将PPT做成视频"、"PPT to video"。 | haoyiyong985 | ↓ 5.6K | ★ 8 |
| 282 | ElevenLabs Speech-to-Text Transcribe audio files using ElevenLabs Speech-to-Text (Scribe v2). | clawdbotborges | ↓ 5.6K | ★ 6 |
| 283 | 蛋糕草图效果图生成-万维易源 面点师通过提示词和蛋糕草图生成蛋糕实际效果图。适用于烘焙设计、蛋糕定制效果图生成、草图还原、视觉丰富化场景 | u_b7ff7d1b | ↓ 5.6K | ★ - |
| 284 | Theme Factory Curated collection of professional color and typography themes for styling artifacts — slides, docs, reports, landing pages. Use when applying visual themes to presentations, generating themed content, or creating custom brand palettes. Triggers on theme, color palette, font pairing, slide styling, presentation theme, brand colors. | wpank | ↓ 5.5K | ★ 3 |
| 285 | Unsplash Search, browse, and download high-quality free photos from Unsplash with filtering, random selection, and detailed photo metadata access. | BrokenWatchCEO | ↓ 5.5K | ★ 3 |
| 286 | MLX STT Speech-To-Text with MLX (Apple Silicon) and opensource models (default GLM-ASR-Nano-2512) locally. | guoqiao | ↓ 5.5K | ★ 2 |
| 287 | Image Generation Create or revise document, PDF, web, or review images with the requested format, sharp raster output, and artifact validation. | xrow GmbH | ↓ 5.5K | ★ 1 |
| 288 | AI Video Restyler Restyle one short video into a new visual treatment while carrying forward the source subject, action, composition, and camera intent. This AI video restyler and video style transfer workflow turns live action into anime, illustration, Chinese comic, ink, clay, paper-cut, or cyberpunk looks from one source clip and a chosen art direction, and reviews style match, subject identity, motion continuity, and source-audio result. Use it for live action to anime, brand visual refresh, Chinese-comic looks, fashion films, music visuals, and creator experiments, with one dominant visual change per run and honest post-result review. | beatra-ai | ↓ 5.5K | ★ 6 |
| 289 | 视频内容智能分析工具 分析视频内容,支持今日头条、抖音、B站、西瓜视频等平台。自动下载视频、语音转文字(faster-whisper)、繁简转换、生成内容摘要。当用户提供视频链接并要求分析视频内容、总结视频、看看视频讲什么时使用此技能。 | user_09d1a78f | ↓ 5.5K | ★ 19 |
| 290 | Transcribe Transcribe audio files to text using local Whisper (Docker). Use when receiving voice messages, audio files (.mp3, .m4a, .ogg, .wav, .webm), or when asked to transcribe audio content. | javicasper | ↓ 5.5K | ★ 2 |
| 291 | Excalidraw Diagram Generator Generate Excalidraw diagrams from natural language descriptions. Use when asked to "create a diagram", "make a flowchart", "visualize a process", "draw a sys... | Morpheous | ↓ 5.5K | ★ 7 |
| 292 | Tts Convert text to speech using Hume AI (or OpenAI) API. Use when the user asks for an audio message, a voice reply, or to hear something "of vive voix". | AMSTKO | ↓ 5.5K | ★ 1 |
| 293 | Canva Connect Manage Canva designs, assets, and folders via the Connect API.
WHAT IT CAN DO:
- List/search/organize designs and folders
- Export finished designs (PNG/PDF/JPG)
- Upload images to asset library
- Autofill brand templates with data
- Create blank designs (doc/presentation/whiteboard/custom)
WHAT IT CANNOT DO:
- Add content to designs (text, shapes, elements)
- Edit existing design content
- Upload documents (images only)
- AI design generation
Best for: asset pipelines, export automation, organization, template autofill.
Triggers: /canva, "upload to canva", "export design", "list my designs", "canva folder". | coolmanns | ↓ 5.5K | ★ 4 |
| 294 | AI歌曲 AI歌曲创作工具,支持一句话写歌、歌词成曲、中文流行歌、品牌歌和歌曲小样生成。 Use this whenever the user wants AI歌曲, ai歌曲, AI歌曲生成, AI唱歌, AI写歌, AI作曲, 生成歌曲, 写一首歌. Call the local aimv CLI through the Node entrypoint. | user_87b8e34f | ↓ 5.5K | ★ 1 |
| 295 | 文本识别OCR 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 | xby-skill | ↓ 5.5K | ★ 1 |
| 296 | AudioPod Use AudioPod AI's API for audio processing tasks including AI music generation (text-to-music, text-to-rap, instrumentals, samples, vocals), stem separation, text-to-speech, noise reduction, speech-to-text transcription, speaker separation, and media extraction. Use when the user needs to generate music/songs/rap from text, split a song into stems/vocals/instruments, generate speech from text, clean up noisy audio, transcribe audio/video, or extract audio from YouTube/URLs. Requires AUDIOPOD_API_KEY env var or pass api_key directly. | Rakesh Roushan | ↓ 5.5K | ★ 5 |
| 297 | Video Analyzer 视频分析处理 — 本地视频反编译分析工具。将视频拆解为时间轴剧本、语音转文字、场景分析、跨模态关联和精华摘要,支持多ASR引擎切换(Whisper/Paraformer/SenseVoice)、中文NLP增强、PaddleOCR中文识别。v4.0 新增短视频平台适配(抖音/快手/B站/视频号)和自动剪辑建议(高光检测/冗余标记/EDL导出/字幕样式)。v4.1 新增tiny模型优先体验(75MB低门槛)、说话人分离质量评分、剪映draft.json导出。v4.2 新增场景管理(detect→slice一条链)、短视频爆款预测、实时直播分析(流式ASR+敏感词检测)。v4.3 新增纯音频输入(mp3/m4a/wav播客与录音)、批量队列(SQLite+硬件档位并发)、GPU自动加速(CT2 int8量化)、ASR配置统一(--asr-engine单参数)。 | fyniujin | ↓ 5.4K | ★ 2 |
| 298 | 即梦-图生视频 CLI 技能,让你的智能体通过 即梦 CLI 生成视频 Use when the user provides image, video, or audio references and wants Dreamina 即梦 video generation through `image2video`, `frames2video`, `multiframe2video`, or `multimodal2video`. Covers CLI v1.4.14 mode routing, required video resolution, Seedance model sets, input limits, and async terminal statuses. | user_c1d16043 | ↓ 5.4K | ★ 4 |
| 299 | 宠物写真-万维易源 基于用户上传的宠物原图,通过视觉特征提取、提示词智能构建与图生图渲染,生成高保真保留宠物本体特征的拟人化高清写真图片;适用于宠物IP形象设计、趣味萌宠拟人化、宠物周边定制场景 | u_38e1d485 | ↓ 5.4K | ★ 1 |
| 300 | Voice Convert text to speech using Microsoft Edge's TTS engine with customizable voices, direct playback, and automatic temporary file cleanup. | zhaov | ↓ 5.4K | ★ - |
Category listings include Claude skills with 5,000+ downloads, merged from Clawhub & SkillHub. See methodology.