@skills · Category
Creative production agent skills
1,836 skills indexed; the 300 most-starred are listed here. Reference any of them in AdaL, Claude Code, Cursor or any coding agent — nothing to install. Search all creative production skills.
- meme-maker · steipete · 385,400 stars
Search meme templates, suggest formats, and generate local or hosted image memes.
- nano-pdf · steipete · 385,400 stars
Edit PDFs with natural-language instructions using the nano-pdf CLI.
- openai-whisper · steipete · 385,400 stars
Local speech-to-text with the Whisper CLI (no API key).
- sag · steipete · 385,400 stars
ElevenLabs text-to-speech with mac-style say UX.
- sherpa-onnx-tts · steipete · 385,400 stars
Local text-to-speech via sherpa-onnx (offline, no cloud)
- brainstorming · obra · 268,212 stars
You MUST use this before any creative work - creating features, building components, adding functionality, or modifying behavior. Explores user intent, req
- fal-ai-media · affaan-m · 238,342 stars
fal.ai MCPによる統合メディア生成(画像、動画、音声)。テキストから画像(Nano Banana)、テキスト/画像から動画(Seedance、Kling、Veo 3)、テキストから音声(CSM-1B)、動画から音声(ThinkSound)をカバーします。ユーザーがAIで画像、動画、音声を生成したい場合
- fal-ai-media · affaan-m · 238,342 stars
通过 fal.ai MCP 实现统一的媒体生成——图像、视频和音频。涵盖文本到图像(Nano Banana)、文本/图像到视频(Seedance、Kling、Veo 3)、文本到语音(CSM-1B),以及视频到音频(ThinkSound)。当用户想要使用 AI 生成图像、视频或音频时使用。
- fal-ai-media · affaan-m · 238,342 stars
Unified media generation via fal.ai MCP — image, video, and audio. Covers text-to-image (Nano Banana), text/image-to-video (Seedance, Kling, Veo 3), text-t
- fal-ai-media · affaan-m · 238,342 stars
Unified media generation via fal.ai MCP — image, video, and audio. Covers text-to-image (Nano Banana), text/image-to-video (Seedance, Kling, Veo 3), text-t
- fal-ai-media · affaan-m · 238,342 stars
Unified media generation via fal.ai MCP — image, video, and audio. Covers text-to-image (Nano Banana), text/image-to-video (Seedance, Kling, Veo 3), text-t
- fal-ai-media · affaan-m · 238,342 stars
Unified media generation via fal.ai MCP — image, video, and audio. Covers text-to-image (Nano Banana), text/image-to-video (Seedance, Kling, Veo 3), text-t
- fal-ai-media · affaan-m · 238,342 stars
通过 fal.ai MCP 实现统一的媒体生成——图像、视频和音频。涵盖文本到图像(Nano Banana)、文本/图像到视频(Seedance、Kling、Veo 3)、文本到语音(CSM-1B),以及视频到音频(ThinkSound)。当用户想要使用 AI 生成图像、视频或音频时使用。
- fal-ai-media · affaan-m · 238,342 stars
fal.ai MCPによる統合メディア生成(画像、動画、音声)。テキストから画像(Nano Banana)、テキスト/画像から動画(Seedance、Kling、Veo 3)、テキストから音声(CSM-1B)、動画から音声(ThinkSound)をカバーします。ユーザーがAIで画像、動画、音声を生成したい場合
- manim-video · affaan-m · 238,342 stars
Build reusable Manim explainers for technical concepts, graphs, system diagrams, and product walkthroughs, then hand off to the wider ECC video stack if ne
- manim-video · affaan-m · 238,342 stars
构建可复用的Manim解释器,用于技术概念、图表、系统图和产品演示,并在需要时移交给更广泛的ECC视频栈。当用户希望获得清晰的动画解释而非通用的人物讲解脚本时使用。
- manim-video · affaan-m · 238,342 stars
构建可复用的Manim解释器,用于技术概念、图表、系统图和产品演示,并在需要时移交给更广泛的ECC视频栈。当用户希望获得清晰的动画解释而非通用的人物讲解脚本时使用。
- manim-video · affaan-m · 238,342 stars
Build reusable Manim explainers for technical concepts, graphs, system diagrams, and product walkthroughs, then hand off to the wider ECC video stack if ne
- nanoclaw-repl · affaan-m · 238,342 stars
操作并扩展NanoClaw v2,这是ECC基于claude -p构建的零依赖会话感知REPL。
- openclaw-persona-forge · affaan-m · 238,342 stars
为 OpenClaw AI Agent 锻造完整的龙虾灵魂方案。根据用户偏好或随机抽卡, 输出身份定位、灵魂描述(SOUL.md)、角色化底线规则、名字和头像生图提示词。 如当前环境提供已审核的生图 skill,可自动生成统一风格头像图片。 当用户需要创建、设计或定制 OpenClaw 龙虾灵魂时使用。 不适
- openclaw-persona-forge · affaan-m · 238,342 stars
为 OpenClaw AI Agent 锻造完整的龙虾灵魂方案。根据用户偏好或随机抽卡, 输出身份定位、灵魂描述(SOUL.md)、角色化底线规则、名字和头像生图提示词。 如当前环境提供已审核的生图 skill,可自动生成统一风格头像图片。 当用户需要创建、设计或定制 OpenClaw 龙虾灵魂时使用。 不适
- openclaw-persona-forge · affaan-m · 238,342 stars
为 OpenClaw AI Agent 锻造完整的龙虾灵魂方案。根据用户偏好或随机抽卡, 输出身份定位、灵魂描述(SOUL.md)、角色化底线规则、名字和头像生图提示词。 如当前环境提供已审核的生图 skill,可自动生成统一风格头像图片。 当用户需要创建、设计或定制 OpenClaw 龙虾灵魂时使用。 不适
- opensource-pipeline · affaan-m · 238,342 stars
Open-source pipeline: fork, sanitize, and package private projects for safe public release. Chains 3 agents (forker, sanitizer, packager). Triggers: '/open
- plan-canvas · affaan-m · 238,342 stars
Open plans and HTML artifacts in a local browser canvas where the human annotates elements, chats, and approves or requests changes without leaving the pag
- plan-orchestrate · affaan-m · 238,342 stars
Read a plan document, decompose it into steps, design a per-step agent chain from the ECC catalogue, and emit ready-to-paste /orchestrate custom prompts. G
- taste · affaan-m · 238,342 stars
A creative-direction (taste) layer for music videos and short-form edits in the angelcore / cloud-trance / hyperpop visual family. Distills a named-genre a
- ui-demo · affaan-m · 238,342 stars
Record polished UI demo videos using Playwright. Use when the user asks to create a demo, walkthrough, screen recording, or tutorial video of a web applica
- ui-demo · affaan-m · 238,342 stars
Record polished UI demo videos using Playwright. Use when the user asks to create a demo, walkthrough, screen recording, or tutorial video of a web applica
- ui-demo · affaan-m · 238,342 stars
使用 Playwright 录制精美的 UI 演示视频。当用户要求创建 Web 应用的演示、导览、屏幕录制或教程视频时使用。生成带有可见光标、自然节奏和专业感的 WebM 视频。
- ui-demo · affaan-m · 238,342 stars
使用 Playwright 录制精美的 UI 演示视频。当用户要求创建 Web 应用的演示、导览、屏幕录制或教程视频时使用。生成带有可见光标、自然节奏和专业感的 WebM 视频。
- video-editing · affaan-m · 238,342 stars
AI-assisted video editing workflows for cutting, structuring, and augmenting real footage. Covers the full pipeline from raw capture through FFmpeg, Remoti
- video-editing · affaan-m · 238,342 stars
実写素材のカット、構築、強化のためのAI支援ビデオ編集ワークフロー。生の撮影素材からFFmpeg、Remotion、ElevenLabs、fal.aiを経て、DescriptまたはCapCutで最終仕上げを行う完全なパイプラインをカバーする。ユーザーがビデオの編集、素材のカット、vlogの作成、またはビデオコ
- video-editing · affaan-m · 238,342 stars
AI辅助的视频编辑工作流程,用于剪辑、构建和增强实拍素材。涵盖从原始拍摄到FFmpeg、Remotion、ElevenLabs、fal.ai,再到Descript或CapCut最终润色的完整流程。适用于用户想要编辑视频、剪辑素材、制作vlog或构建视频内容的情况。
- video-editing · affaan-m · 238,342 stars
AI辅助的视频编辑工作流程,用于剪辑、构建和增强实拍素材。涵盖从原始拍摄到FFmpeg、Remotion、ElevenLabs、fal.ai,再到Descript或CapCut最终润色的完整流程。适用于用户想要编辑视频、剪辑素材、制作vlog或构建视频内容的情况。
- video-editing · affaan-m · 238,342 stars
AI-assisted video editing workflows for cutting, structuring, and augmenting real footage. Covers the full pipeline from raw capture through FFmpeg, Remoti
- video-editing · affaan-m · 238,342 stars
実写素材のカット、構築、強化のためのAI支援ビデオ編集ワークフロー。生の撮影素材からFFmpeg、Remotion、ElevenLabs、fal.aiを経て、DescriptまたはCapCutで最終仕上げを行う完全なパイプラインをカバーする。ユーザーがビデオの編集、素材のカット、vlogの作成、またはビデオコ
- video-editing · affaan-m · 238,342 stars
AI-assisted video editing workflows for cutting, structuring, and augmenting real footage. Covers the full pipeline from raw capture through FFmpeg, Remoti
- video-editing · affaan-m · 238,342 stars
AI-assisted video editing workflows for cutting, structuring, and augmenting real footage. Covers the full pipeline from raw capture through FFmpeg, Remoti
- videodb · affaan-m · 238,342 stars
视频与音频的查看、理解与行动。查看:从本地文件、URL、RTSP/直播源或实时录制桌面获取内容;返回实时上下文和可播放流链接。理解:提取帧,构建视觉/语义/时间索引,并通过时间戳和自动剪辑搜索片段。行动:转码和标准化(编解码器、帧率、分辨率、宽高比),执行时间线编辑(字幕、文本/图像叠加、品牌化、音频叠加、配
- videodb · affaan-m · 238,342 stars
ビデオとオーディオの表示、理解、アクション。表示:ローカルファイル、URL、RTSP/ライブストリーム、またはリアルタイムのデスクトップ録画からコンテンツを取得し、リアルタイムコンテキストと再生可能なストリームリンクを返す。理解:フレームを抽出し、ビジュアル/セマンティック/時間的インデックスを構築し、タイム
- videodb · affaan-m · 238,342 stars
视频与音频的查看、理解与行动。查看:从本地文件、URL、RTSP/直播源或实时录制桌面获取内容;返回实时上下文和可播放流链接。理解:提取帧,构建视觉/语义/时间索引,并通过时间戳和自动剪辑搜索片段。行动:转码和标准化(编解码器、帧率、分辨率、宽高比),执行时间线编辑(字幕、文本/图像叠加、品牌化、音频叠加、配
- videodb · affaan-m · 238,342 stars
ビデオとオーディオの表示、理解、アクション。表示:ローカルファイル、URL、RTSP/ライブストリーム、またはリアルタイムのデスクトップ録画からコンテンツを取得し、リアルタイムコンテキストと再生可能なストリームリンクを返す。理解:フレームを抽出し、ビジュアル/セマンティック/時間的インデックスを構築し、タイム
- visa-doc-translate · affaan-m · 238,342 stars
Translate visa application documents (images) to English and create a bilingual PDF with original and translation. Use when visa application document image
- visa-doc-translate · affaan-m · 238,342 stars
Translate visa application documents (images) to English and create a bilingual PDF with original and translation. Use when visa application document image
- ascii-video · nousresearch · 226,679 stars
ASCII video: convert video/audio to colored ASCII MP4/GIF.
- audiocraft-audio-generation · nousresearch · 226,679 stars
AudioCraft: MusicGen text-to-music, AudioGen text-to-sound.
- baoyu-article-illustrator · nousresearch · 226,679 stars
Article illustrations: type × style × palette consistency.
- baoyu-comic · nousresearch · 226,679 stars
Knowledge comics (知识漫画): educational, biography, tutorial.
- creative-ideation · nousresearch · 226,679 stars
Generate ideas via named methods from creative practice.
- google_meet · nousresearch · 226,679 stars
Join a Google Meet call, transcribe live captions, optionally speak in realtime, and do the followup work afterwards. Use when the user asks the agent to s
- heartmula · nousresearch · 226,679 stars
HeartMuLa: Suno-like song generation from lyrics + tags.
- hyperframes · nousresearch · 226,679 stars
Render MP4/WebM videos from HTML compositions.
- manim-video · nousresearch · 226,679 stars
Manim CE animations: 3Blue1Brown math/algo videos.
- meme-generation · nousresearch · 226,679 stars
Create meme PNGs from templates with Pillow text overlay.
- p5js · nousresearch · 226,679 stars
p5.js sketches: gen art, shaders, interactive, 3D.
- pdf · nousresearch · 226,679 stars
PDF files: create, read, merge, fill, OCR, edit text.
- pixel-art · nousresearch · 226,679 stars
Pixel art w/ era palettes (NES, Game Boy, PICO-8).
- powerpoint · nousresearch · 226,679 stars
Create, read, edit .pptx decks with python-pptx.
- songsee · nousresearch · 226,679 stars
Audio spectrograms/features (mel, chroma, MFCC) via CLI.
- songwriting-and-ai-music · nousresearch · 226,679 stars
Songwriting craft and Suno AI music prompts.
- stable-diffusion · nousresearch · 226,679 stars
Text-to-image generation, inpainting, and img2img.
- tldraw-offline · nousresearch · 226,679 stars
Drive and script tldraw offline canvases with an agent.
- whisper · nousresearch · 226,679 stars
- algorithmic-art · anthropics · 166,745 stars
Creating algorithmic art using p5.js with seeded randomness and interactive parameter exploration. Use this when users request creating art using code, gen
- pdf · anthropics · 166,745 stars
Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multi
- pptx · anthropics · 166,745 stars
Use this skill any time a .pptx or .potx file is involved in any way — as input, output, or both. This includes: creating slide decks, pitch decks, or pres
- slack-gif-creator · anthropics · 166,745 stars
Knowledge and utilities for creating animated GIFs optimized for Slack. Provides constraints, validation tools, and animation concepts. Use when users requ
- wowerpoint · thedotmack · 89,894 stars
Turn one document into a kawaii NotebookLM slide-deck PDF. Use for "wowerpoint this", "make a deck about <file>", "turn this report into slides", or any re
- ai-music-album · nexu-io · 84,236 stars
Full-lifecycle AI music album production — concept, lyric drafting, track sequencing, and export. Useful for indie album experiments and brand soundtracks.
- algorithmic-art · nexu-io · 84,236 stars
Create generative art using p5.js with seeded randomness so every render is reproducible. Useful for procedural posters, motion-style stills, and artistic
- audio-jingle · nexu-io · 84,236 stars
Audio generation skill — jingles, beds, voiceover, and sound effects. Routes music requests to Suno V5 / Udio / Lyria, speech to MiniMax TTS / FishAudio /
- audio-jingle · nexu-io · 84,236 stars
Audio generation skill — jingles, beds, voiceover, and sound effects. Routes music requests to Suno V5 / Udio / Lyria, speech to MiniMax TTS / FishAudio /
- brainstorming · nexu-io · 84,236 stars
Transform rough ideas into fully-formed designs through structured questioning and alternative exploration. Useful early in concept work.
- chat-motion-overlay · nexu-io · 84,236 stars
Generate configurable chat motion overlays from a transcript or screenshot, including plain bubble scenes, app-style chat containers, optional device frame
- create-hyperframes-launch · nexu-io · 84,236 stars
Use this plugin when the user wants a HyperFrames-ready HTML motion composition, launch animation, kinetic typography clip, product reveal, or social video
- create-video-storyboard · nexu-io · 84,236 stars
Use this plugin when the user wants a video concept, storyboard, shot list, prompt pack, or render-ready motion brief for a product, campaign, or explainer
- creative-director · nexu-io · 84,236 stars
AI creative director with recursive self-assessment: 20+ methodologies (SIT, TRIZ, Bisociation, SCAMPER, Synectics), 3-axis evaluation calibrated against C
- critique-theater · nexu-io · 84,236 stars
Five-dimension design quality review — score the artifact against craft, brand, accessibility, and copy, then fix what falls short before handing it over.
- fal-3d · nexu-io · 84,236 stars
Generate 3D models from text or images via fal.ai. Useful for game assets, AR previews, product mockups, and concept sculpting.
- fal-generate · nexu-io · 84,236 stars
Generate images and videos using fal.ai AI models. Production-grade catalogue covering Flux, SDXL, ideogram, and other community-hosted endpoints.
- fal-image-edit · nexu-io · 84,236 stars
AI-powered image editing with style transfer, background removal, object removal, and inpainting via fal.ai hosted models.
- fal-kling-o3 · nexu-io · 84,236 stars
Generate images and videos with Kling O3 — Kling's most powerful model family — via fal.ai.
- fal-lip-sync · nexu-io · 84,236 stars
Create talking head videos and lip sync audio to video via fal.ai. Useful for explainer avatars, multilingual dubbing previews, and social cuts.
- fal-realtime · nexu-io · 84,236 stars
Real-time and streaming AI image generation via fal.ai. Suited for moodboard exploration, draft variations, and rapid creative iteration.
- fal-restore · nexu-io · 84,236 stars
Restore and fix image quality — deblur, denoise, fix faces, and restore old documents using fal.ai's hosted restoration models.
- fal-upscale · nexu-io · 84,236 stars
Upscale and enhance image and video resolution using AI super-resolution models hosted on fal.ai.
- fal-video-edit · nexu-io · 84,236 stars
Edit existing videos using AI — remix style, upscale, remove background, and add audio via fal.ai's hosted video models.
- gif-sticker-maker · nexu-io · 84,236 stars
Convert photos into animated GIF stickers in Funko Pop / Pop Mart style via the MiniMax API. Useful for personalized chat stickers and avatar packs.
- hps-memphis-pop · nexu-io · 84,236 stars
A pop-culture retrospective on how 1980s design language shaped today's apps — the scenes, the turning point, and the takeaway. Built as a decision-grade s
- humanize-ppt · nexu-io · 84,236 stars
A presentation system for agent-made PPTs — born for the talk, not just the template. It turns raw material into an AST (audience-state-transfer) outline w
- hyperframes · nexu-io · 84,236 stars
Create video compositions, animations, title cards, overlays, captions, voiceovers, audio-reactive visuals, and scene transitions in HyperFrames HTML. Use
- hyperframes · nexu-io · 84,236 stars
Create video compositions, animations, title cards, overlays, captions, voiceovers, audio-reactive visuals, and scene transitions in HyperFrames HTML. Use
- library-curator · nexu-io · 84,236 stars
Search the OD Library (the global asset registry) and apply matching assets into the current project mid-task. Use when the user asks to reuse an image the
- nanobanana-ppt · nexu-io · 84,236 stars
AI-powered PPT generation with document analysis and styled images via the NanoBanana stack. Combines image generation with structured deck output.
- od-media-generation · nexu-io · 84,236 stars
Default reference pipeline for image, video, and audio projects — routes through media-image / media-video / media-audio atoms based on the project kind, w
- pptx · nexu-io · 84,236 stars
Read, generate, and adjust PowerPoint slides, layouts, and templates. Useful for executive decks, training material, and product reviews.
- pptx-generator · nexu-io · 84,236 stars
Create and edit PowerPoint presentations from scratch with PptxGenJS — MiniMax's production-tested deck pipeline.
- remotion · nexu-io · 84,236 stars
Programmatic video creation with React. Useful for branded explainers, social cuts, dashboards-to-video, and reproducible motion graphics.
- slack-gif-creator · nexu-io · 84,236 stars
Create animated GIFs optimized for Slack with validators for size constraints and composable animation primitives.
- slides · nexu-io · 84,236 stars
Create and edit .pptx presentation decks with PptxGenJS. Useful for sales decks, kickoff briefs, and design-system showcases.
- sora · nexu-io · 84,236 stars
Generate, remix, and manage short video clips via OpenAI's Sora API. Useful for cinematic shots, b-roll, and rapid concept video iteration.
- speech · nexu-io · 84,236 stars
Generate spoken audio from text using OpenAI's API with built-in voices. Useful for narrated explainers, lecture audio, and quick voiceover tracks.
- ve-terminal-mono · nexu-io · 84,236 stars
OpenDesign from the CLI: driving the full design workflow with the `od` command — scripted, composable, agent-ready. Built as a decision-grade AI literacy
- venice-audio-music · nexu-io · 84,236 stars
Music generation queueing, retrieval, and completion endpoints via Venice.ai. Suited for jingles, background loops, and prototype scoring.
- venice-audio-speech · nexu-io · 84,236 stars
Text-to-speech models, voices, formats, and streaming via Venice.ai. Useful for narration, voiceover, and conversational agent voices.
- venice-image-edit · nexu-io · 84,236 stars
Image edits, upscaling, and background removal via the Venice.ai API.
- venice-image-generate · nexu-io · 84,236 stars
Image generation endpoints and available styles via the Venice.ai API.
- venice-video · nexu-io · 84,236 stars
Video generation and transcription workflows via the Venice.ai API.
- video-downloader · nexu-io · 84,236 stars
Download videos from YouTube and other platforms for offline viewing, editing, or archival with support for various formats and quality options.
- video-hyperframes · nexu-io · 84,236 stars
Hyperframes / Remotion-compatible continuous frame animation with autoplay support.
- video-hyperframes · nexu-io · 84,236 stars
Hyperframes / Remotion-compatible continuous frame animation with autoplay support.
- video-shortform · nexu-io · 84,236 stars
Short-form video generation skill — 3-10 second clips for product reveals, motion teasers, ambient loops. Defaults to Seedance 2 but works the same with Kl
- video-shortform · nexu-io · 84,236 stars
Short-form video generation skill — 3-10 second clips for product reveals, motion teasers, ambient loops. Defaults to Seedance 2 but works the same with Kl
- video-template-frame-build-minimal · nexu-io · 84,236 stars
Use this plugin when the user wants a "Build Minimal Frame" HyperFrames motion video — Luxury-minimal whitespace hero — single word reveals letter by lette
- video-template-frame-creative-voltage · nexu-io · 84,236 stars
Use this plugin when the user wants a "Creative Voltage Frame" HyperFrames motion video — Electric split with hand-drawn script — offset panels slide in, d
- video-template-frame-data-chart-nyt · nexu-io · 84,236 stars
Use this plugin when the user wants a "NYT-Style Data Chart Frame" HyperFrames motion video — NYT-newsroom typography, staggered reveal animation, and edit
- video-template-frame-nyt-graph · nexu-io · 84,236 stars
Use this plugin when the user wants a "NYT Graph" HyperFrames motion video — Animated data chart in print editorial style
- video-template-frame-play-mode · nexu-io · 84,236 stars
Use this plugin when the user wants a "Play Mode" HyperFrames motion video — Playful elastic animations
- video-template-frame-swiss-grid · nexu-io · 84,236 stars
Use this plugin when the user wants a "Swiss Grid" HyperFrames motion video — Structured grid layout
- video-template-frame-takram-organic · nexu-io · 84,236 stars
Use this plugin when the user wants a "Takram Organic Frame" HyperFrames motion video — Soft-tech radial node graph as art — frosted rounded card, curved l
- video-template-vfx-text-cursor · nexu-io · 84,236 stars
Use this plugin when the user wants a "VFX Text Cursor" HyperFrames motion video — Cursor light trail, chromatic rays, and directional flares for word-by-w
- youtube-clipper · nexu-io · 84,236 stars
YouTube clip generation and editing with automated workflows — pull source video, slice highlights, add captions, and export.
- music-generation · bytedance · 79,444 stars
Use this skill when the user requests to generate, create, compose, or produce music or songs — background music, theme songs, jingles, or instrumental tra
- podcast-generation · bytedance · 79,444 stars
Use this skill when the user requests to generate, create, or produce podcasts from text content. Converts written content into a two-host conversational p
- ppt-generation · bytedance · 79,444 stars
Use this skill when the user requests to generate, create, or make presentations (PPT/PPTX). Creates visually rich slides by generating images for each sli
- surprise-me · bytedance · 79,444 stars
Create a delightful, unexpected "wow" experience for the user by dynamically discovering and creatively combining other enabled skills. Triggers when the u
- video-generation · bytedance · 79,444 stars
Use this skill when the user requests to generate, create, or imagine videos. Supports structured prompts and reference image for guided generation.
- googleslides-automation · composiohq · 71,973 stars
Automate Google Slides tasks via Rube MCP (Composio): create presentations, add slides from Markdown, batch update, copy from templates, get thumbnails. Al
- HeyGen Automation · composiohq · 71,973 stars
Automate AI video generation, avatar browsing, template-based video creation, and video status tracking through HeyGen's platform via Composio
- pptx · composiohq · 71,973 stars
Presentation creation, editing, and analysis. When Claude needs to work with presentations (.pptx files) for: (1) Creating new presentations, (2) Modifying
- slack-gif-creator · composiohq · 71,973 stars
Toolkit for creating animated GIFs optimized for Slack, with validators for size constraints and composable animation primitives. This skill applies when u
- remotion-render · remotion-dev · 55,698 stars
Export a Remotion video
- slidev · slidevjs · 47,992 stars
Create and present web-based slidedecks for developers using Slidev with Markdown, Vue components, code highlighting, animations, and interactive features.
- cli-anything-audacity · hkuds · 46,739 stars
Command-line interface for Audacity - A stateful command-line interface for audio editing, following the same patterns as the GIMP and Ble...
- cli-anything-drawio · hkuds · 46,739 stars
Command-line interface for Drawio - A CLI harness for **Draw.io** — create, edit, and export diagrams from the command line....
- cli-anything-gimp · hkuds · 46,739 stars
Command-line interface for Gimp - A stateful command-line interface for image editing, built on Pillow. Designed for AI agents and pow...
- cli-anything-kdenlive · hkuds · 46,739 stars
Command-line interface for Kdenlive - A stateful command-line interface for video editing, following the same patterns as the Blender CLI ...
- cli-anything-krita · hkuds · 46,739 stars
CLI harness for Krita digital painting — manage projects, layers, filters, and export via command line. Use when automating Krita workflows, batch processi
- cli-anything-openscreen · hkuds · 46,739 stars
Command-line interface for Openscreen — a screen recording editor. A stateful CLI for editing screen recordings with zoom, speed ramps, trim, crop, annotat
- cli-anything-openscreen · hkuds · 46,739 stars
Command-line interface for Openscreen — a screen recording editor. A stateful CLI for editing screen recordings with zoom, speed ramps, trim, crop, annotat
- cli-anything-videocaptioner · hkuds · 46,739 stars
AI-powered video captioning — transcribe speech, optimize/translate subtitles, and burn them into video via the stable VideoCaptioner backend. Free ASR and
- cli-anything-videocaptioner · hkuds · 46,739 stars
AI-powered video captioning — transcribe speech, optimize/translate subtitles, and burn them into video via the stable VideoCaptioner backend. Free ASR and
- cli-hub-matrix-video-creation · hkuds · 46,739 stars
Capability-based multi-tool matrix for video production. Agents pick providers (CLI-Anything harnesses, public CLIs, Python libs, native binaries, cloud AP
- musescore · hkuds · 46,739 stars
CLI for music notation — transpose, export PDF/audio/MIDI, extract parts, manage instruments
- algorithmic-art · sickn33 · 44,571 stars
Algorithmic philosophies are computational aesthetic movements that are then expressed through code. Output .md files (philosophy), .html files (interactiv
- algorithmic-art · sickn33 · 44,571 stars
Algorithmic philosophies are computational aesthetic movements that are then expressed through code. Output .md files (philosophy), .html files (interactiv
- algorithmic-art · sickn33 · 44,571 stars
Algorithmic philosophies are computational aesthetic movements that are then expressed through code. Output .md files (philosophy), .html files (interactiv
- algorithmic-art · sickn33 · 44,571 stars
Algorithmic philosophies are computational aesthetic movements that are then expressed through code. Output .md files (philosophy), .html files (interactiv
- algorithmic-art · sickn33 · 44,571 stars
Algorithmic philosophies are computational aesthetic movements that are then expressed through code. Output .md files (philosophy), .html files (interactiv
- article-illustrations · sickn33 · 44,571 stars
Generate hand-drawn 16:9 article illustrations with the Grav character IP, sparse annotations, and absurd but clear visual metaphors.
- article-illustrations · sickn33 · 44,571 stars
Generate hand-drawn 16:9 article illustrations with the Grav character IP, sparse annotations, and absurd but clear visual metaphors.
- article-illustrations · sickn33 · 44,571 stars
Generate hand-drawn 16:9 article illustrations with the Grav character IP, sparse annotations, and absurd but clear visual metaphors.
- daily-gift · sickn33 · 44,571 stars
Relationship-aware daily gift engine with five-stage creative pipeline — editorial judgment, synthesis, concept generation, visual strategy, and rendering
- daily-gift · sickn33 · 44,571 stars
Relationship-aware daily gift engine with five-stage creative pipeline — editorial judgment, synthesis, concept generation, visual strategy, and rendering
- daily-gift · sickn33 · 44,571 stars
Relationship-aware daily gift engine with five-stage creative pipeline — editorial judgment, synthesis, concept generation, visual strategy, and rendering
- diary · sickn33 · 44,571 stars
Unified Diary System: A context-preserving automated logger for multi-project development.
- fal-audio · sickn33 · 44,571 stars
Text-to-speech and speech-to-text using fal.ai audio models
- fal-audio · sickn33 · 44,571 stars
Text-to-speech and speech-to-text using fal.ai audio models
- fal-generate · sickn33 · 44,571 stars
Generate images and videos using fal.ai AI models
- fal-generate · sickn33 · 44,571 stars
Generate images and videos using fal.ai AI models
- fal-generate · sickn33 · 44,571 stars
Generate images and videos using fal.ai AI models
- fal-upscale · sickn33 · 44,571 stars
Upscale and enhance image and video resolution using AI
- fal-upscale · sickn33 · 44,571 stars
Upscale and enhance image and video resolution using AI
- fal-upscale · sickn33 · 44,571 stars
Upscale and enhance image and video resolution using AI
- gemini-omni-flash-api · sickn33 · 44,571 stars
Use this skill for generative video editing, text-to-video, image-referenced video generation, and first-frame-to-video transition animations using the off
- gemini-omni-flash-api · sickn33 · 44,571 stars
Use this skill for generative video editing, text-to-video, image-referenced video generation, and first-frame-to-video transition animations using the off
- gemini-omni-flash-api · sickn33 · 44,571 stars
Use this skill for generative video editing, text-to-video, image-referenced video generation, and first-frame-to-video transition animations using the off
- generate-nanobanana · sickn33 · 44,571 stars
Generate and edit images/video with Google's Gemini media models (Nano Banana 2/Pro, Gemini Omni Flash), with cost-approval gates, reference-image support,
- generate-nanobanana · sickn33 · 44,571 stars
Generate and edit images/video with Google's Gemini media models (Nano Banana 2/Pro, Gemini Omni Flash), with cost-approval gates, reference-image support,
- generate-nanobanana · sickn33 · 44,571 stars
Generate and edit images/video with Google's Gemini media models (Nano Banana 2/Pro, Gemini Omni Flash), with cost-approval gates, reference-image support,
- image-generator · sickn33 · 44,571 stars
Generate and edit images using Gemini's Nano Banana Pro model (gemini-3-pro-image-preview). Use this skill when the user asks you to generate images, creat
- image-generator · sickn33 · 44,571 stars
Generate and edit images using Gemini's Nano Banana Pro model (gemini-3-pro-image-preview). Use this skill when the user asks you to generate images, creat
- image-generator · sickn33 · 44,571 stars
Generate and edit images using Gemini's Nano Banana Pro model (gemini-3-pro-image-preview). Use this skill when the user asks you to generate images, creat
- image-studio · sickn33 · 44,571 stars
Studio de geracao de imagens inteligente — roteamento automatico entre ai-studio-image (fotos humanizadas/influencer) e stability-ai (arte/ ilustracao/edic
- image-studio · sickn33 · 44,571 stars
Studio de geracao de imagens inteligente — roteamento automatico entre ai-studio-image (fotos humanizadas/influencer) e stability-ai (arte/ ilustracao/edic
- image-studio · sickn33 · 44,571 stars
Studio de geracao de imagens inteligente — roteamento automatico entre ai-studio-image (fotos humanizadas/influencer) e stability-ai (arte/ ilustracao/edic
- impress · sickn33 · 44,571 stars
Presentation creation, format conversion (ODP/PPTX/PDF), slide automation with LibreOffice Impress.
- impress · sickn33 · 44,571 stars
Presentation creation, format conversion (ODP/PPTX/PDF), slide automation with LibreOffice Impress.
- impress · sickn33 · 44,571 stars
Presentation creation, format conversion (ODP/PPTX/PDF), slide automation with LibreOffice Impress.
- llm-council · sickn33 · 44,571 stars
Run Fireworks-hosted open-weight model councils that compare responses and synthesize a final answer.
- loop-library · sickn33 · 44,571 stars
Find, compare, adapt, and design bounded AI-agent feedback loops with explicit checks, stop rules, guardrails, and handoffs.
- macos-screen-recorder · sickn33 · 44,571 stars
macOS screen recorder that captures the main display PLUS system audio via ScreenCaptureKit — no BlackHole/loopback driver, no sudo, just the standard Scre
- mmx-cli · sickn33 · 44,571 stars
Use mmx to generate text, images, video, speech, and music via the MiniMax AI platform. Use when the user wants to create media content, chat with MiniMax
- mmx-cli · sickn33 · 44,571 stars
Use mmx to generate text, images, video, speech, and music via the MiniMax AI platform. Use when the user wants to create media content, chat with MiniMax
- mmx-cli · sickn33 · 44,571 stars
Use mmx to generate text, images, video, speech, and music via the MiniMax AI platform. Use when the user wants to create media content, chat with MiniMax
- modellix · sickn33 · 44,571 stars
Integrate the Modellix API/CLI for async AI image, video, and speech generation or transcription (model run --wait, task download).
- modellix · sickn33 · 44,571 stars
Integrate the Modellix API/CLI for async AI image, video, and speech generation or transcription (model run --wait, task download).
- modellix · sickn33 · 44,571 stars
Integrate the Modellix API/CLI for async AI image, video, and speech generation or transcription (model run --wait, task download).
- nanobanana-ppt-skills · sickn33 · 44,571 stars
AI-powered PPT generation with document analysis and styled images
- podcast-generation · sickn33 · 44,571 stars
Generate real audio narratives from text content using Azure OpenAI's Realtime API.
- podcast-generation · sickn33 · 44,571 stars
Generate real audio narratives from text content using Azure OpenAI's Realtime API.
- podcast-generation · sickn33 · 44,571 stars
Generate real audio narratives from text content using Azure OpenAI's Realtime API.
- pptx-official · sickn33 · 44,571 stars
A user may ask you to create, edit, or analyze the contents of a .pptx file. A .pptx file is essentially a ZIP archive containing XML files and other resou
- pptx-official · sickn33 · 44,571 stars
A user may ask you to create, edit, or analyze the contents of a .pptx file. A .pptx file is essentially a ZIP archive containing XML files and other resou
- pptx-official · sickn33 · 44,571 stars
A user may ask you to create, edit, or analyze the contents of a .pptx file. A .pptx file is essentially a ZIP archive containing XML files and other resou
- pptx-official · sickn33 · 44,571 stars
A user may ask you to create, edit, or analyze the contents of a .pptx file. A .pptx file is essentially a ZIP archive containing XML files and other resou
- pptx-official · sickn33 · 44,571 stars
A user may ask you to create, edit, or analyze the contents of a .pptx file. A .pptx file is essentially a ZIP archive containing XML files and other resou
- remotion · sickn33 · 44,571 stars
Generate walkthrough videos from Stitch projects using Remotion with smooth transitions, zooming, and text overlays
- remotion · sickn33 · 44,571 stars
Generate walkthrough videos from Stitch projects using Remotion with smooth transitions, zooming, and text overlays
- remotion · sickn33 · 44,571 stars
Generate walkthrough videos from Stitch projects using Remotion with smooth transitions, zooming, and text overlays
- screenstudio-alt · sickn33 · 44,571 stars
Open-source headless Screen Studio alternative: auto speed-up of idle, auto-zoom on click clusters, keystroke overlay chips, smoothed synthetic cursor, and
- slack-gif-creator · sickn33 · 44,571 stars
A toolkit providing utilities and knowledge for creating animated GIFs optimized for Slack.
- speed · sickn33 · 44,571 stars
Launch RSVP speed reader for text
- speed · sickn33 · 44,571 stars
Launch RSVP speed reader for text
- stability-ai · sickn33 · 44,571 stars
Geracao de imagens via Stability AI (SD3.5, Ultra, Core). Text-to-image, img2img, inpainting, upscale, remove-bg, search-replace. 15 estilos artisticos.
- stability-ai · sickn33 · 44,571 stars
Geracao de imagens via Stability AI (SD3.5, Ultra, Core). Text-to-image, img2img, inpainting, upscale, remove-bg, search-replace. 15 estilos artisticos.
- stability-ai · sickn33 · 44,571 stars
Geracao de imagens via Stability AI (SD3.5, Ultra, Core). Text-to-image, img2img, inpainting, upscale, remove-bg, search-replace. 15 estilos artisticos.
- steve-jobs · sickn33 · 44,571 stars
Agente que simula Steve Jobs — cofundador da Apple, CEO da Pixar, fundador da NeXT, o maior designer de produtos tecnologicos da historia e o mais influent
- steve-jobs · sickn33 · 44,571 stars
Agente que simula Steve Jobs — cofundador da Apple, CEO da Pixar, fundador da NeXT, o maior designer de produtos tecnologicos da historia e o mais influent
- video-content-extractor · sickn33 · 44,571 stars
Extract key frames from MP4 videos at configurable intervals, run Tesseract OCR, and generate structured Markdown reports with video metadata and timestamp
- video-content-extractor · sickn33 · 44,571 stars
Extract key frames from MP4 videos at configurable intervals, run Tesseract OCR, and generate structured Markdown reports with video metadata and timestamp
- video-content-extractor · sickn33 · 44,571 stars
Extract key frames from MP4 videos at configurable intervals, run Tesseract OCR, and generate structured Markdown reports with video metadata and timestamp
- videodb · sickn33 · 44,571 stars
Video and audio perception, indexing, and editing. Ingest files/URLs/live streams, build visual/spoken indexes, search with timestamps, edit timelines, add
- videodb · sickn33 · 44,571 stars
Video and audio perception, indexing, and editing. Ingest files/URLs/live streams, build visual/spoken indexes, search with timestamps, edit timelines, add
- videodb · sickn33 · 44,571 stars
Video and audio perception, indexing, and editing. Ingest files/URLs/live streams, build visual/spoken indexes, search with timestamps, edit timelines, add
- videodb-skills · sickn33 · 44,571 stars
Upload, stream, search, edit, transcribe, and generate AI video and audio using the VideoDB SDK.
- videodb-skills · sickn33 · 44,571 stars
Upload, stream, search, edit, transcribe, and generate AI video and audio using the VideoDB SDK.
- videodb-skills · sickn33 · 44,571 stars
Upload, stream, search, edit, transcribe, and generate AI video and audio using the VideoDB SDK.
- voice-agents · sickn33 · 44,571 stars
Voice agents represent the frontier of AI interaction - humans
- voice-agents · sickn33 · 44,571 stars
Voice agents represent the frontier of AI interaction - humans
- web-media-getter · sickn33 · 44,571 stars
One query across free image / video / GIF APIs (stock + historical/archival + GIF engines), returning normalized, license-tagged results with optional top-
- web-media-getter · sickn33 · 44,571 stars
One query across free image / video / GIF APIs (stock + historical/archival + GIF engines), returning normalized, license-tagged results with optional top-
- youtube-notetaker · sickn33 · 44,571 stars
Turn YouTube talks into local study notes with slides, transcripts, editable annotations, and a markdown-backed viewer.
- youtube-notetaker · sickn33 · 44,571 stars
Turn YouTube talks into local study notes with slides, transcripts, editable annotations, and a markdown-backed viewer.
- youtube-notetaker · sickn33 · 44,571 stars
Turn YouTube talks into local study notes with slides, transcripts, editable annotations, and a markdown-backed viewer.
- ppt-master · hugohe3 · 43,577 stars
AI-driven presentation workflow for generating editable PPTX decks and slides, reconstructing page visuals, creating reusable Brand/Style/Layout/Deck works
- video · coreyhaines31 · 43,366 stars
When the user wants to create, generate, or produce video content using AI tools or programmatic frameworks. Also use when the user mentions 'video product
- captions-overlay · heygen-com · 39,818 stars
Overlay doctrine for the embedded-captions workflow — the caption MODEL (drop / rail / embed) and the rule that captions are an OVERLAY composited on top o
- captions-overlay · heygen-com · 39,818 stars
Overlay doctrine for the embedded-captions workflow — the caption MODEL (drop / rail / embed) and the rule that captions are an OVERLAY composited on top o
- changelog-video · heygen-com · 39,818 stars
Turn a weekly changelog .md into a finished branded changelog video (square 1080, ~45-60s, Annie VO, animated brand background, mock-UI visualizations, low
- changelog-video · heygen-com · 39,818 stars
Turn a weekly changelog .md into a finished branded changelog video (square 1080, ~45-60s, Annie VO, animated brand background, mock-UI visualizations, low
- cut-the-curve · heygen-com · 39,818 stars
The technique catalog: five velocity-matched SEAMS (zoom-through, INVERSE zoom-through, cut-the-curve, waterfall cut, rack-focus blur-cut) plus the two in-
- embedded-captions · heygen-com · 39,818 stars
Add captions or subtitles to an existing single-subject talking-head video without editing the footage. Use for plain verbatim captions, cinematic captions
- faceless-explainer · heygen-com · 39,818 stars
Turn arbitrary text — an article, notes, a topic, a brief — into a faceless explainer video: there is no site or footage to capture, so the visuals are inv
- general-video · heygen-com · 39,818 stars
Author or edit a custom HyperFrames composition when no specialized workflow fits, or when BRIEF.md sets flow: companion. Use for longer or multi-scene pie
- hyperframes · heygen-com · 39,818 stars
Mandatory entry point: read this first for any request to make, create, edit, animate, or render a video, animation, or motion graphic, including a promo,
- media-use · heygen-com · 39,818 stars
Agent Media OS, the single skill for every media need in a HyperFrames project. Resolve BGM, SFX, image, icon, brand logo, voice, color grade, or LUT into
- music-to-video · heygen-com · 39,818 stars
Turn a music track (an audio file, a video to pull audio from, or a track generated from a mood brief) into a beat-synced video — lyric video, slideshow, o
- pr-to-video · heygen-com · 39,818 stars
Turn a GitHub pull request (a PR URL, owner/repo#N, or 'this PR' in a checked-out repo) into a code-change explainer video — changelog, feature reveal, fix
- product-launch-video · heygen-com · 39,818 stars
Turn a product or marketing URL, pasted script, or brief into a product launch / promo video — SaaS promos, feature reveals, product demos, app and company
- talking-head-recut · heygen-com · 39,818 stars
Package an existing talking-head / interview / podcast video with timed, designed GRAPHIC OVERLAY cards — kinetic titles, lower-thirds, data callouts, quot
- file-conversion · wshobson · 38,570 stars
Convert files between formats — PDF to Word, HEIC to JPG, MP4 to MP3, CSV to JSON, EPUB to MOBI, and 999 total routes across images, video, audio, document
- agentic-eval · github · 37,534 stars
Patterns and techniques for evaluating and improving AI agent outputs. Use this skill when: - Implementing self-critique and reflection loops - Building ev
- convert-pdf-to-md · github · 37,534 stars
Converts PDF (.pdf) documents into Markdown so their contents can be accurately analyzed, summarized, searched, or extracted from. Use this skill whenever
- game-engine · github · 37,534 stars
Expert skill for building web-based game engines and games using HTML5, Canvas, WebGL, and JavaScript. Use when asked to create games, build game engines,
- image-manipulation-image-magick · github · 37,534 stars
Process and manipulate images using ImageMagick. Supports resizing, format conversion, batch processing, and retrieving image metadata. Use when working wi
- latchshot-page-capture · github · 37,534 stars
Use this skill when a user needs a screenshot, website thumbnail, full-page capture, or PDF of a public HTTP(S) webpage saved as a local artifact through L
- nano-banana-pro-openrouter · github · 37,534 stars
Generate or edit images via OpenRouter with the Gemini 3 Pro Image model. Use for prompt-only image generation, image edits, and multi-image compositing; s
- plantuml-ascii · github · 37,534 stars
Generate ASCII art diagrams using PlantUML text mode. Use when user asks to create ASCII diagrams, text-based diagrams, terminal-friendly diagrams, or ment
- screen-recording · github · 37,534 stars
Create annotated animated GIF demos and screen recordings for pull requests and documentation. Covers frame capture, timing, imageio-based GIF creation, an
- transloadit-media-processing · github · 37,534 stars
Process media files (video, audio, images, documents) using Transloadit. Use when asked to encode video to HLS/MP4, generate thumbnails, resize or watermar
- ppt-template-creator · anthropics · 34,059 stars
Creates self-contained PPT template SKILLS (not presentations) from user-provided PowerPoint templates. Use ONLY when a user wants to create a reusable ski
- pptx-author · anthropics · 34,059 stars
Produce a .pptx file on disk (headless) instead of driving a live PowerPoint document — for managed-agent sessions with no open Office app.
- pptx-author · anthropics · 34,059 stars
Produce a .pptx file on disk (headless) instead of driving a live PowerPoint document — for managed-agent sessions with no open Office app.
- pptx-author · anthropics · 34,059 stars
Produce a .pptx file on disk (headless) instead of driving a live PowerPoint document — for managed-agent sessions with no open Office app.
- pptx-author · anthropics · 34,059 stars
Produce a .pptx file on disk (headless) instead of driving a live PowerPoint document — for managed-agent sessions with no open Office app.
- nature-figure · yuan1z0825 · 33,818 stars
Create, revise, audit, and export submission-grade scientific figures for Nature-family and other high-impact venues in Python (matplotlib/seaborn) or R (g
- nature-paper2ppt · yuan1z0825 · 33,818 stars
Build a complete Nature-style Chinese PPTX presentation from a scientific paper, preprint, PDF, article text, figure legends, or reading notes. Use for jou
- matplotlib · k-dense-ai · 32,849 stars
Low-level plotting library for full customization. Use when you need fine-grained control over every plot element, creating novel plot types, or integratin
- pdf · k-dense-ai · 32,849 stars
Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multi
- pptx · k-dense-ai · 32,849 stars
Use this skill any time a .pptx or .potx file is involved in any way — as input, output, or both. This includes: creating slide decks, pitch decks, or pres
- pptx-posters · k-dense-ai · 32,849 stars
Create and audit editable scientific posters in macro-free PowerPoint (.pptx) from author-approved local content and assets. Use when the requested deliver
- scientific-slides · k-dense-ai · 32,849 stars
Build slide decks and presentations for research talks. Use this for making PowerPoint slides, conference presentations, seminar talks, research presentati
- recipe-create-doc-from-template · googleworkspace · 30,245 stars
Copy a Google Docs template, fill in content, and share with collaborators.
- recipe-create-presentation · googleworkspace · 30,245 stars
Create a new Google Slides presentation and add initial slides.
- algorithmic-art · davila7 · 30,138 stars
Creating algorithmic art using p5.js with seeded randomness and interactive parameter exploration. Use this when users request creating art using code, gen
- audiocraft-audio-generation · davila7 · 30,138 stars
PyTorch library for audio generation including text-to-music (MusicGen) and text-to-sound (AudioGen). Use when you need to generate music from text descrip
- brainstorming · davila7 · 30,138 stars
You MUST use this before any creative work - creating features, building components, adding functionality, or modifying behavior. Explores user intent, req
- brainstorming · davila7 · 30,138 stars
You MUST use this before any creative work - creating features, building components, adding functionality, or modifying behavior. Explores user intent, req
- luma-imagegen · davila7 · 30,138 stars
Use when the user asks to generate images via the Luma AI API (Dream Machine / Photon); collects a prompt and options interactively, then calls the API usi
- manim · davila7 · 30,138 stars
Comprehensive guide for Manim Community - Python framework for creating mathematical animations and educational videos with programmatic control
- meme-factory · davila7 · 30,138 stars
Generate memes using the memegen.link API. Use when users request memes, want to add humor to content, or need visual aids for social media. Supports 100+
- pdf · davila7 · 30,138 stars
Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multi
- pdf · davila7 · 30,138 stars
Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multi
- pdf · davila7 · 30,138 stars
Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multi
- pdf-fill-studio · davila7 · 30,138 stars
Fill any PDF locally and place each value precisely in a visual editor. Use when the user wants to fill out a PDF form, enter data into a PDF, complete a t
- pdf-official · davila7 · 30,138 stars
Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multi
- remotion · davila7 · 30,138 stars
Best practices and comprehensive guide for Remotion - programmatic video creation in React with animations, compositions, and media handling
- scientific-slides · davila7 · 30,138 stars
Build slide decks and presentations for research talks. Use this for making PowerPoint slides, conference presentations, seminar talks, research presentati
- screenshot · davila7 · 30,138 stars
Use when the user explicitly asks for a desktop or system screenshot (full screen, specific app or window, or a pixel region), or when tool-specific captur
- slack-gif-creator · davila7 · 30,138 stars
Knowledge and utilities for creating animated GIFs optimized for Slack. Provides constraints, validation tools, and animation concepts. Use when users requ
- speech · davila7 · 30,138 stars
Use when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation. OpenAI remains the defaul
- stable-diffusion-image-generation · davila7 · 30,138 stars
State-of-the-art text-to-image generation with Stable Diffusion models via HuggingFace Diffusers. Use when generating images from text prompts, performing
- video-downloader · davila7 · 30,138 stars
Downloads videos from YouTube and other platforms for offline viewing, editing, or archival. Handles various formats and quality options.
- whisper · davila7 · 30,138 stars
OpenAI's general-purpose speech recognition model. Supports 99 languages, transcription, translation to English, and language identification. Six model siz
- baoyu-comic · jimliu · 24,677 stars
Knowledge comic creator supporting multiple art styles and tones. Creates original educational comics with detailed panel layouts and batch-capable image g
- baoyu-compress-image · jimliu · 24,677 stars
Compresses images to WebP (default) or PNG with automatic tool selection. Use when user asks to "compress image", "optimize image", "convert to webp", or r
- baoyu-danger-gemini-web · jimliu · 24,677 stars
Generates images and text via reverse-engineered Gemini Web API. Supports text generation, image generation from prompts, reference images for vision input
- baoyu-image-gen · jimliu · 24,677 stars
AI image generation with OpenAI GPT Image 2.5, Azure OpenAI, Google, OpenRouter, DashScope, Z.AI GLM-Image, MiniMax, Jimeng, Seedream, Replicate and Agnes
- baoyu-markdown-to-html · jimliu · 24,677 stars
Converts Markdown to styled HTML with WeChat-compatible themes. Supports code highlighting, math, Mermaid (rendered to PNG via headless Chrome), PlantUML,
- baoyu-post-to-weibo · jimliu · 24,677 stars
Posts content to Weibo (微博). Supports regular posts with text, images, and videos, and headline articles (头条文章) with Markdown input via Chrome CDP. Use whe
- baoyu-slide-deck · jimliu · 24,677 stars
Generates professional slide deck images from content. Creates outlines with style instructions, then generates individual slide images. Use when user asks
- screenshot · openai · 24,594 stars
Use when the user explicitly asks for a desktop or system screenshot (full screen, specific app or window, or a pixel region), or when tool-specific captur
- speech · openai · 24,594 stars
Use when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API;
- demo-video · alirezarezvani · 23,939 stars
Use when the user asks to create a demo video, product walkthrough, feature showcase, animated presentation, marketing video, or GIF from screenshots or sc
- demo-video · alirezarezvani · 23,939 stars
- md-review · alirezarezvani · 23,939 stars
Converts a markdown PR writeup or code review (one with ```diff fenced blocks and severity-tagged > [!BLOCKER]/[!MAJOR]/[!MINOR]/[!NIT] callouts) into a si
- view-pdf · anthropics · 23,345 stars
Interactive PDF viewer. Use when the user wants to open, show, or view a PDF and collaborate on it visually — annotate, highlight, stamp, fill form fields,
- watch · bradautomates · 14,305 stars
Watch a video (URL or local path). Downloads with yt-dlp, extracts auto-scaled frames with ffmpeg, pulls the transcript from captions (or Whisper API fallb
- buddy-sings · minimax-ai · 13,260 stars
Use when user wants their Claude Code pet (/buddy) to sing a song. Triggers on any request that combines the concept of their Claude Code buddy, pet, or co