Media Generation AI Agent Skills

Explore 1,892 Media Generation AI agent skills on AgentCC. Compare source repositories, technical metadata, and installation guidance.

nano-banana-pro

246.8k
openclawopenclaw

Generate or edit images via Gemini 3 Pro Image (Nano Banana Pro).

137 days ago

canvas-design

81.1k
anthropicsanthropics

Create beautiful visual art in .png and .pdf documents using design philosophy. You should use this skill when the user asks to create a poster, piece of art, design, or other static piece. Create original visual designs, never copying existing artists' work to avoid copyright violations.

137 days ago

content-research-writer

39.6k
ComposioHQComposioHQ

Assists in writing high-quality content by conducting research, adding citations, improving hooks, iterating on outlines, and providing real-time feedback on each section. Transforms your writing process from solo effort to collaborative partnership.

137 days ago

openai-image-gen

246.8k
openclawopenclaw

Batch-generate images via OpenAI Images API. Random prompt sampler + `index.html` gallery.

137 days ago

algorithmic-art

81.1k
anthropicsanthropics

Creating algorithmic art using p5.js with seeded randomness and interactive parameter exploration. Use this when users request creating art using code, generative art, algorithmic art, flow fields, or particle systems. Create original algorithmic art rather than copying existing artists' work to avoid copyright violations.

137 days ago

slack-gif-creator

39.6k
ComposioHQComposioHQ

Toolkit for creating animated GIFs optimized for Slack, with validators for size constraints and composable animation primitives. This skill applies when users request animated GIFs or emoji animations for Slack from descriptions like "make me a GIF for Slack of X doing Y".

137 days ago

songsee

246.8k
openclawopenclaw

Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.

137 days ago

slack-gif-creator

81.1k
anthropicsanthropics

Knowledge and utilities for creating animated GIFs optimized for Slack. Provides constraints, validation tools, and animation concepts. Use when users request animated GIFs for Slack like "make me a GIF of X doing Y for Slack."

137 days ago

image-enhancer

39.6k
ComposioHQComposioHQ

Improves the quality of images, especially screenshots, by enhancing resolution, sharpness, and clarity. Perfect for preparing images for presentations, documentation, or social media posts.

137 days ago

video-frames

246.8k
openclawopenclaw

Extract frames or short clips from videos using ffmpeg.

137 days ago

brand-guidelines

81.1k
anthropicsanthropics

Applies Anthropic's official brand colors and typography to any sort of artifact that may benefit from having Anthropic's look-and-feel. Use it when brand colors or style guidelines, visual formatting, or company design standards apply.

137 days ago

latex-posters

21.8k
davila7davila7

Create professional research posters in LaTeX using beamerposter, tikzposter, or baposter. Support for conference presentations, academic posters, and scientific communication. Includes layout design, color schemes, multi-column formats, figure integration, and poster-specific best practices for visual communication.

137 days ago

gifgrep

246.8k
openclawopenclaw

Search GIF providers with CLI/TUI, download results, and extract stills/sheets.

137 days ago

last30days

18.0k
sickn33sickn33

Research a topic from the last 30 days on Reddit + X + Web, become an expert, and write copy-paste-ready prompts for the user's target tool.

137 days ago

stable-diffusion-image-generation

21.8k
davila7davila7

State-of-the-art text-to-image generation with Stable Diffusion models via HuggingFace Diffusers. Use when generating images from text prompts, performing image-to-image translation, inpainting, or building custom diffusion pipelines.

Image GenerationStable DiffusionDiffusers+3
137 days ago

sherpa-onnx-tts

246.8k
openclawopenclaw

Local text-to-speech via sherpa-onnx (offline, no cloud)

137 days ago

fal-image-edit

18.0k
sickn33sickn33

AI-powered image editing with style transfer and object removal

137 days ago

viral-generator-builder

21.8k
davila7davila7

Expert in building shareable generator tools that go viral - name generators, quiz makers, avatar creators, personality tests, and calculator tools. Covers the psychology of sharing, viral mechanics, and building tools people can't resist sharing with friends. Use when: generator tool, quiz maker, name generator, avatar creator, viral tool.

137 days ago

image-generation

17.6k
onyx-dot-apponyx-dot-app

Generate images using nano banana.

137 days ago

Frontend Slides

18.0k
sickn33sickn33

Create stunning, animation-rich HTML presentations from scratch or by converting PowerPoint files. Use when the user wants to build a presentation, convert a PPT/PPTX to web, or create slides for a talk, tutorial, or update.

137 days ago

meme-factory

21.8k
davila7davila7

Generate memes using the memegen.link API. Use when users request memes, want to add humor to content, or need visual aids for social media. Supports 100+ popular templates with custom text and styling.

137 days ago

generate-image

10.8k
K-Dense-AIK-Dense-AI

Generate or edit images using AI models (FLUX, Nano Banana 2). Use for general-purpose image generation including photos, illustrations, artwork, visual assets, concept art, and any image that is not a technical diagram or schematic. For flowcharts, circuits, pathways, and technical diagrams, use the scientific-schematics skill instead.

137 days ago

screenshots

18.0k
sickn33sickn33

Generate marketing screenshots of your app using Playwright. Use when the user wants to create screenshots for Product Hunt, social media, landing pages, or documentation.

137 days ago

generate-image

21.8k
davila7davila7

Generate or edit images using AI models (FLUX, Gemini). Use for general-purpose image generation including photos, illustrations, artwork, visual assets, concept art, and any image that isn't a technical diagram or schematic. For flowcharts, circuits, pathways, and technical diagrams, use the scientific-schematics skill instead.

137 days ago

pptx-posters

10.8k
K-Dense-AIK-Dense-AI

Create research posters using HTML/CSS that can be exported to PDF or PPTX. Use this skill ONLY when the user explicitly requests PowerPoint/PPTX poster format. For standard research posters, use latex-posters instead. This skill provides modern web-based poster design with responsive layouts and easy visual integration.

137 days ago

seo-content-writer

18.0k
sickn33sickn33

Writes SEO-optimized content based on provided keywords and topic briefs. Creates engaging, comprehensive content following best practices. Use PROACTIVELY for content creation tasks.

137 days ago

content-creator

21.8k
davila7davila7

Create SEO-optimized marketing content with consistent brand voice. Includes brand voice analyzer, SEO optimizer, content frameworks, and social media templates. Use when writing blog posts, creating social media content, analyzing brand voice, optimizing SEO, planning content calendars, or when user mentions content creation, brand voice, SEO optimization, social media marketing, or content strategy.

137 days ago

scientific-schematics

10.8k
K-Dense-AIK-Dense-AI

Create publication-quality scientific diagrams using Nano Banana Pro AI with smart iterative refinement. Uses Gemini 3 Pro for quality review. Only regenerates if quality is below threshold for your document type. Specialized in neural network architectures, system diagrams, flowcharts, biological pathways, and complex scientific visualizations.

137 days ago

beautiful-prose

18.0k
sickn33sickn33

Hard-edged writing style contract for timeless, forceful English prose without AI tics

137 days ago

audiocraft-audio-generation

21.8k
davila7davila7

PyTorch library for audio generation including text-to-music (MusicGen) and text-to-sound (AudioGen). Use when you need to generate music from text descriptions, create sound effects, or perform melody-conditioned music generation.

MultimodalAudio GenerationText-to-Music+2
137 days ago

paper-2-web

10.8k
K-Dense-AIK-Dense-AI

This skill should be used when converting academic papers into promotional and presentation formats including interactive websites (Paper2Web), presentation videos (Paper2Video), and conference posters (Paper2Poster). Use this skill for tasks involving paper dissemination, conference preparation, creating explorable academic homepages, generating video abstracts, or producing print-ready posters from LaTeX or PDF sources.

137 days ago

fal-audio

18.0k
sickn33sickn33

Text-to-speech and speech-to-text using fal.ai audio models

137 days ago

humanizer

21.8k
davila7davila7

Remove signs of AI-generated writing from text. Use when editing or reviewing text to make it sound more natural and human-written. Based on Wikipedia's comprehensive "Signs of AI writing" guide. Detects and fixes patterns including: inflated symbolism, promotional language, superficial -ing analyses, vague attributions, em dash overuse, rule of three, AI vocabulary words, negative parallelisms, and excessive conjunctive phrases. Credits: Original skill by @blader - https://github.com/blader/humanizer

137 days ago

infographics

10.8k
K-Dense-AIK-Dense-AI

Create professional infographics using Nano Banana Pro AI with smart iterative refinement. Uses Gemini 3 Pro for quality review. Integrates research-lookup and web search for accurate data. Supports 10 infographic types, 8 industry styles, and colorblind-safe palettes.

137 days ago

sora

10.4k
openaiopenai

Use when the user asks to generate, remix, poll, list, download, or delete Sora videos via OpenAI’s video API using the bundled CLI (`scripts/sora.py`), including requests like “generate AI video,” “Sora,” “video remix,” “download video/thumbnail/spritesheet,” and batch video generation; requires `OPENAI_API_KEY` and Sora API access.

137 days ago

paper-2-web

21.8k
davila7davila7

This skill should be used when converting academic papers into promotional and presentation formats including interactive websites (Paper2Web), presentation videos (Paper2Video), and conference posters (Paper2Poster). Use this skill for tasks involving paper dissemination, conference preparation, creating explorable academic homepages, generating video abstracts, or producing print-ready posters from LaTeX or PDF sources.

137 days ago

blog-post

9.8k
langchain-ailangchain-ai

Use this skill when writing long-form blog posts, tutorials, or educational articles that require structure, depth, and SEO considerations

137 days ago

speech

10.4k
openaiopenai

Use when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API; run the bundled CLI (`scripts/text_to_speech.py`) with built-in voices and require `OPENAI_API_KEY` for live calls. Custom voice creation is out of scope.

137 days ago

gemini-imagegen

9.7k
EveryIncEveryInc

This skill should be used when generating and editing images using the Gemini API (Nano Banana Pro). It applies when creating images from text prompts, editing existing images, applying style transfers, generating logos with text, creating stickers, product mockups, or any image generation/manipulation task. Supports text-to-image, image editing, multi-turn refinement, and composition from multiple reference images.

137 days ago

content-creation

8.4k
anthropicsanthropics

Draft marketing content across channels — blog posts, social media, email newsletters, landing pages, press releases, and case studies. Use when writing any marketing content, when you need channel-specific formatting, SEO-optimized copy, headline options, or calls to action.

137 days ago

imagegen

10.4k
openaiopenai

Use when the user asks to generate or edit images via the OpenAI Image API (for example: generate image, edit/inpaint/mask, background removal or replacement, transparent background, product shots, concept art, covers, or batch variants); run the bundled CLI (`scripts/image_gen.py`) and require `OPENAI_API_KEY` for live calls.

137 days ago

Baoyu — XHS Images

6.1k
JimLiuJimLiu

Generates Xiaohongshu (Little Red Book) infographic series with 10 visual styles and 8 layouts. Splits content into 1–10 cartoon-style images optimized for Xiaohongshu/XHS engagement. Use when a user mentions "Xiaohongshu images", "XHS images", "RedNote infographics", "Xiaohongshu seeding", or requests social-media infographics for Chinese platforms.

137 days ago

infographic-creator

4.5k
antvisantvis

Create beautiful infographics based on the given text content. Used when the user requests to create an infographic.

137 days ago

humanizer-en

2.9k
op7418op7418

Remove traces of AI generation from text. Suitable for editing or reviewing text to make it sound more natural and more like human writing. Based on Wikipedia's comprehensive guide to 'AI Writing Features'. Detect and fix the following patterns: exaggerated symbolism, propagandistic language, shallow analysis ending in -ing, vague attribution, overuse of dashes, three-part rule, AI vocabulary, negative parallelism, excessive connective phrases.

137 days ago

baoyu-article-illustrator

6.1k
JimLiuJimLiu

Analyzes article structure, identifies positions requiring visual aids, generates illustrations with Type × Style two-dimensional approach. Use when user asks to "illustrate article", "add images", "generate images for article", or "illustrate article".

137 days ago

muapi-media-generation

2.8k
SamurAIGPTSamurAIGPT

Generate AI images, videos, music, and audio from the terminal via muapi.ai — supports 100+ models including Flux, Midjourney v7, Kling 3.0, Veo3, and Suno V5

137 days ago

brand-guidelines

2.5k
davepoondavepoon

Applies Anthropic's official brand colors and typography to any sort of artifact that may benefit from having Anthropic's look-and-feel. Use it when brand colors or style guidelines, visual formatting, or company design standards apply.

137 days ago

baoyu-danger-gemini-web

6.1k
JimLiuJimLiu

Generates images and text via reverse-engineered Gemini Web API. Supports text generation, image generation from prompts, reference images for vision input, and multi-turn conversations. Use when other skills need image generation backend, or when user requests 'generate image with Gemini', 'Gemini text generation', or needs vision-capable AI generation.

137 days ago