Skills

All Skills

image

Skills tagged with #image

@utkarsh

nextjs-seo-2026

Complete SEO and Answer Engine Optimization (AEO) skill for Next.js 15 apps. Covers metadata, structured data (JSON-LD), sitemap, robots.txt, Core Web Vitals, dynamic OG images, and AI crawler optimization. Use this skill when building or auditing SEO for any Next.js project.

SEOSearch Engine Optimization
5mo ago
1
@SlavaSexton

krea

Use when a task names Krea or Krea 2, when choosing between Krea's hosted API and its open weights, when wiring the Krea 2 Image or Krea 2 Style Reference API nodes, when building with FLUX.1 Krea Dev, when someone asks about Krea Realtime or realtime or streaming video generation, or when a Krea 2 job needs ControlNet, instruction editing, identity preservation or per-layer conditioning control that core ComfyUI does not provide.

SlavaSexton/ComfyUI-Agent-Kit+1 more
1mo ago
670
@mcp-registry
MCP

NSFW Image Detector

NSFWJS-based image safety detector (GIF/APNG/WebP/JPEG/PNG).

mcpgithubweb
5mo ago
0
@Ascend

CANN Image Release Engineer

You are a software development engineer responsible for automating and executing the CANN Docker image release pipeline.

Ascend/cann-container-image
5mo ago
430
@Whitemarmot
MCP

Io.Github.Whitemarmot/Markly

Add text or logo watermarks to images via the Markly.cloud API. Batch supported.

mcpgithubapi
Whitemarmot/markly-mcp-server
5mo ago
0
@Railly

ray

Generate code screenshots via ray.tinte.dev API. Use when user asks for code screenshots, code images, code snippets as images, or says "ray". Calls POST https://ray.tinte.dev/api/v1/screenshot and saves the PNG result.

Railly/tinte+1 more
5mo ago
5720
@erajasekar
MCP

AI Diagram Maker

Generate software diagrams from natural language, code, ASCII, images, Mermaid.

mcpgithubai
erajasekar/ai-diagram-maker-mcp
5mo ago
0
@ethanjyx

openbrand

Extract brand assets (logos, colors, backdrop images, brand name) from any website URL. Use when building branded interfaces, generating style guides, or needing brand identity data from a URL.

ethanjyx/OpenBrand
5mo ago
1990
@AgriciDaniel

blog

Full-lifecycle blog engine with 17 commands, 12 content templates, 5-category 100-point scoring, and 4 specialized agents. Optimized for Google rankings (December 2025 Core Update, E-E-A-T) and AI citations (GEO/AEO). Writes, rewrites, analyzes, outlines, audits, and repurposes blog content with answer-first formatting, sourced statistics, Pixabay/Unsplash/Pexels images, AI image generation via Gemini, built-in SVG chart generation, JSON-LD schema generation, and freshness signals. Supports any platform (WordPress, Next.js MDX, Hugo, Ghost, Astro, Jekyll, 11ty, Gatsby, HTML). Use when user says "blog", "write blog", "blog post", "blog strategy", "content brief", "editorial calendar", "analyze blog", "rewrite blog", "update blog", "blog SEO", "blog optimization", "content plan", "blog outline", "seo check", "schema markup", "repurpose", "geo audit", "blog audit", "citation readiness".

AgriciDaniel/claude-blog+20 more
5mo ago
2300
@jqknono
MCP

Pic Gen

Unified image generation tool - cover images (cover) and Mermaid diagrams (mermaid) generation

mcpgithubai
jqknono/pic-gen
5mo ago
0
@prism-php

developing-with-prism

Guide for developing with Prism PHP package - a Laravel package for integrating LLMs. Activate or use when working with Prism features including text generation, structured output, embeddings, image generation, audio processing, streaming, tools/function calling, or any LLM provider integration (OpenAI, Anthropic, Gemini, Mistral, Groq, XAI, DeepSeek, OpenRouter, Ollama, VoyageAI, ElevenLabs). Activate for any Prism-related development tasks.

prism-php/prism
5mo ago
2.3K0
@whale-professor
MCP

Io.Github.Whale Professor/Mcp Server 3daistudio

MCP server for 3D AI Studio — generate 3D models from text or images using Hunyuan and TRELLIS

mcpgithubai
whale-professor/3daistudio_skill
5mo ago
0
@mcp-registry
MCP

Mcp Server

AI agent tools: web search, browser, 400+ LLMs, image gen, TTS, phone verify. Pay-per-use.

mcpgithubapiaisearchbrowser
5mo ago
0
@Pwd9000-ML

new-blog-post

Scaffold, draft, and validate a new DEV.to blog post. Use when: creating a new article, starting a new blog post, scaffolding a post, generating a cover image, checking post quality, validating front matter, fixing post formatting, or auditing an existing post.

Pwd9000-ML/blog-devto
5mo ago
300
@michtio

craft-cloud

Craft Cloud — Pixel & Tonic's serverless hosting platform for Craft CMS. Covers craft-cloud.yaml configuration, the Build → Migrate → Release deploy pipeline, the craftcms/cloud extension package, edge image transforms via Cloudflare, edge static caching with cache.rules + ESI, Cloud-managed S3 filesystem, MySQL 8 / Postgres 15 databases (no MariaDB, no tablePrefix), Console-based command runner and scheduled cron (once-per-hour minimum), auto-handled queue jobs, custom domains and SSL, preview environments per branch, Cloud limitations (ephemeral filesystem, no SSH, no .htaccess, no built-in mail), plugin development requirements for Cloud compatibility, and self-hosted → Cloud migration. Triggers on: craft-cloud.yaml, craftcms/cloud package, cloud.esi(), php craft cloud/up, php craft cloud/setup, App::isEphemeral(), CRAFT_EPHEMERAL, edge.craft.cloud, preview.craft.cloud, CRAFT_CLOUD_PROJECT_ID, CRAFT_CLOUD_ENVIRONMENT_ID, CRAFT_CLOUD_CDN_BASE_URL, Build → Migrate → Release, Cloud filesystem, Cloud-compatible plugin, Cloudflare Images at edge, AssetsFs, static-caching rules, ESI islands, deploy to Craft Cloud, migrate to Craft Cloud, self-hosted to Cloud, Craft Cloud quotas, Craft Cloud regions. Do NOT trigger for generic Craft deployment (Forge, Servd, bare metal) — that lives in craftcms/deployment.md. Do NOT trigger for general DDEV local dev unrelated to Cloud parity.

michtio/craftcms-claude-skills+8 more
4mo ago
460
@Santazuki

Unblind

Route images to vision API. Never pretend to see. Never Read/Edit settings.json.

Santazuki/unblind
4mo ago
120
@Libres-coder
MCP

Io.Github.Libres Coder/Parseflow

PDF parsing server with text extraction, metadata, search, images, and TOC via MCP

mcpgithubsearch
Libres-coder/ParseFlow
5mo ago
0
@AgriciDaniel

banana

AI image generation Creative Director powered by Google Gemini Nano Banana models. Use this skill for ANY request involving image creation, editing, visual asset production, or creative direction. Triggers on: generate an image, create a photo, edit this picture, design a logo, make a banner, visual for my anything, and all /banana commands. Handles text-to-image, image editing, multi-turn creative sessions, batch workflows, and brand presets.

AgriciDaniel/banana-claude
5mo ago
620
@poco-ai

MiniMax Multi-Modal Toolkit

Generate voice, music, video, and image content via MiniMax APIs — the unified entry for **MiniMax multimodal** use cases (audio + music + video + image). Includes voice cloning & voice design for custom voices, image generation with character reference, and FFmpeg-based media tools for audio/vide

poco-ai/poco-claw
5mo ago
1.3K0
@hyperlane-xyz

build-docker-image

Trigger Docker image builds for Hyperlane agent, monorepo, or node service images. Use when the user wants to build new Docker images for a branch, commit, or tag.

hyperlane-xyz/hyperlane-monorepo+14 more
5mo ago
520
@AceDataCloud
MCP

Io.Github.AceDataCloud/Mcp Midjourney

MCP server for Midjourney AI image generation and editing

mcpgithubai
AceDataCloud/MCPMidjourney
5mo ago
0
@guanyang

baoyu-compress-image

Compresses images to WebP (default) or PNG with automatic tool selection. Use when user asks to "compress image", "optimize image", "convert to webp", or reduce image file size.

guanyang/antigravity-skills+21 more
4mo ago
7610
@amilich

isometric-asset-sheets

Generate sprite sheet images for the isometric city game using the GenerateImage tool. Use when creating new game assets, sprite sheets, vehicle sprites, building sprites, or any visual assets for the isometric city builder. Ensures consistent format, sizing, and isometric projections.

amilich/isometric-city+1 more
5mo ago
2.1K0
@OguntolaIbrahim
MCP

Io.Github.OguntolaIbrahim/Image Viewer

InChat Image Viewer MCP - View images inline in AI chat by just providing the file path

mcpgithubaifile
OguntolaIbrahim/inchat-image-viewer-mcp
5mo ago
0
@jezweb

ai-image-generator

Generate AI images using Gemini or GPT APIs directly. Covers model selection (Gemini for scenes, GPT for transparent icons), the 5-part prompting framework, API calling patterns, multi-turn editing, and quality assurance. Produces photorealistic scenes, icons, illustrations, OG images, and product shots. Use when building websites that need images, creating marketing assets, or generating visual content. Triggers: 'generate image', 'ai image', 'create hero image', 'make an icon', 'generate illustration', 'create og image', 'ai art', 'image generation'.

jezweb/claude-skills+55 more
5mo ago
6320
@sidart10
MCP

Runway AI Video Generation

AI video generation with Gen-4, Veo 3, and Aleph editing. Text-to-video, image-to-video, 4K upscale

mcpgithubai
sidart10/runway-mcp-server
5mo ago
0
@cloudinary
MCP

Cloudinary Asset Management

Upload, organize, search, and transform images, videos, and files with AI-powered tools.

mcpgithubaisearchfile
cloudinary/asset-management-mcp
5mo ago
0
@pijusz
MCP

Mcp Sanity Images

MCP server for uploading local images to Sanity CMS

mcpgithub
pijusz/mcp-sanity-images
5mo ago
0
@LarryWalkerDEV
MCP

ImmoStage Virtual Staging

AI virtual staging for real estate — stage rooms, beautify floor plans, classify images.

mcpgithubai
LarryWalkerDEV/mcp-immostage
5mo ago
0
@AceDataCloud
MCP

Io.Github.AceDataCloud/Mcp Nanobanana Pro

MCP server for NanoBanana AI image generation and editing

mcpgithubai
AceDataCloud/MCPNanoBanana
5mo ago
0
@rounak

phoneagent

Control a connected iPhone, iOS simulator, Android emulator, or Android device from macOS through PhoneAgent's JSON-RPC bridge. Use when users ask to automate mobile UI actions, inspect accessibility trees, toggle Settings switches, navigate apps, or capture screenshots by sending RPC methods like get_tree, get_screen_image, get_context, tap_element, enter_text, scroll, swipe, and open_app.

rounak/PhoneAgent
5mo ago
7260
@jau123

Creative Toolkit

Generate professional AI images through a unified interface that routes across multiple providers. Search curated prompts, enhance ideas into production-ready descriptions, and manage local ComfyUI workflows — all from a single MCP server.

jau123/MeiGen-AI-Design-MCP+2 more
5mo ago
5380
@smixs

Image Prompting — Nano Banana & GPT Image 2

This skill writes image prompts. It does not generate images. The output is: model name + quality / size / aspect ratio + the prompt itself.

smixs/visual-skills
5mo ago
210
@isfendipgensin

creatomate

Creatomate video/image generation API. Use when: (1) Creating videos or images programmatically with the Creatomate SDK, (2) Building slideshows, concatenating videos, adding text overlays, (3) Adding animated captions or subtitles, (4) Generating social media content (TikTok, Instagram Stories, YouTube Shorts), (5) Using templates with dynamic modifications, (6) Applying effects (blur, filters, masks, transitions), (7) Integrating with ChatGPT for AI-generated content, (8) Working with compositions and animations.

isfendipgensin/claude-code-skill-creatomate
5mo ago
10
@OmidZamani

dspy-adapters-multimodal

This skill should be used when the user asks to "choose a DSPy adapter", "use JSONAdapter", "use XMLAdapter", "enable native function calling", "send images, audio, or files to DSPy", mentions `dspy.ChatAdapter`, `dspy.JSONAdapter`, `dspy.XMLAdapter`, `dspy.Image`, `dspy.Audio`, `dspy.File`, structured outputs, or multimodal DSPy signatures.

OmidZamani/dspy-skills+5 more
4mo ago
780
@selftune-dev

ai-image-generation

Facilitates creation of visual content using generative AI models including DALL-E, Midjourney prompts, and Stable Diffusion workflows.

selftune-dev/selftune+1 more
5mo ago
70
@shamspias

fennec-image-compression

Use this skill when asked to compress, resize, or analyze images in Go using the Fennec library, or when modifying the Fennec codebase itself.

shamspias/fennec
5mo ago
640
@TinyAGI

imagegen

Use when the user asks to generate or edit images via the OpenAI Image API (for example: generate image, edit/inpaint/mask, background removal or replacement, transparent background, product shots, concept art, covers, or batch variants); run the bundled CLI (`scripts/image_gen.py`) and require `OPENAI_API_KEY` for live calls.

TinyAGI/tinyclaw+4 more
5mo ago
3.1K0
@securecoders
MCP

OpenGraph.io MCP Server

MCP server for OpenGraph.io API - fetch OG data, screenshots, scrape, and generate images

mcpgithubapi
securecoders/opengraph-io-mcp
5mo ago
0
@alonw0

web-asset-generator

Generate web assets including favicons, app icons (PWA), and social media meta images (Open Graph) for Facebook, Twitter, WhatsApp, and LinkedIn. Use when users need icons, favicons, social sharing images, or Open Graph images from logos or text slogans. Handles image resizing, text-to-image generation, and provides proper HTML meta tags.

alonw0/web-asset-generator
5mo ago
2090
@kgelster

shopify-alt-text

Use when backfilling missing image alt text on a Shopify store: images with no alt attribute, accessibility alt text for product images, alt text for Files library images, image SEO, or a WCAG/accessibility pass that flags images missing text alternatives. Triggers: "backfill alt text", "fill in missing alt text", "images have no alt", "add alt attributes for accessibility", "image SEO sweep". One mutation, fileUpdate, covers both product media and the Files library. Not for meta titles/descriptions (use shopify-seo-metadata).

kgelster/awesome-ecom-skills+5 more
3mo ago
90
@instavm

image-crop-rotate

Image processing skill for cropping images to 50% from center and rotating them 90 degrees clockwise. This skill should be used when users request image cropping to center, image rotation, or both operations combined on image files.

instavm/coderunner+1 more
5mo ago
8010
@rediumvex

/social-captions — Algorithm-Optimized Social Media Captions (v2)

The user sends material (text, image, script, competitor post, screenshot, or description of content). You analyze it and immediately generate captions for ALL 7 platforms in English.

rediumvex/social-media-caption-generator-claude
5mo ago
410
@kdeldycke

rename-with-dates

Rename documents and files (PDFs, images, screenshots, etc.) by reading their content to extract the effective/publication date, then renaming them with a "YYYY-MM-DD - Clear descriptive title.ext" format. Use when the user wants to organize files with date prefixes based on document content.

kdeldycke/dotfiles
5mo ago
1650
@Positronic-Robotics

remote-training

Manages remote training infrastructure on Nebius VMs. Use for building/pushing Docker images, starting/stopping VM machines (train, train2, train3), running training jobs, dataset generation, and starting inference servers.

Positronic-Robotics/positronic
5mo ago
620
@nexscope-ai

amazon-a-plus-content

Plan and create Amazon A+ Content (Enhanced Brand Content). Design module layouts, write persuasive copy, plan comparison charts, and create image briefs that convert browsers into buyers.

nexscope-ai/Amazon-Skills+35 more
5mo ago
610
@imjuya

juya-news-card-operator

Operate juya-news-card as a local CLI text-to-image card tool. Use when Codex should transform user text into final card content itself and render PNG directly with Playwright plus SSR-ready templates, without starting the Next.js app.

imjuya/juya-news-card
5mo ago
750
@srod

node-minify

Compress JavaScript, CSS, HTML, JSON, and image files using node-minify library. Use when: minifying/compressing assets, bundling JS/CSS files, optimizing images (WebP/AVIF), concatenating files, or when user mentions "node-minify", "@node-minify", "minification". Triggers: "minify", "compress JS/CSS", "bundle", "optimize images", "reduce file size".

srod/node-minify
5mo ago
5160
@PyJudge
MCP

Io.Github.PyJudge/Pdf4vllm

PDF reader for vision LLMs. Auto-detects text corruption and switches to image mode.

mcpgithubllm
PyJudge/pdf4vllm-mcp
5mo ago
0
@ChrBoebel
MCP

Optical Context MCP

Compress OCR-heavy PDFs into dense packed images so agents can work with long visual documents.

mcpgithub
ChrBoebel/optical-context-mcp
5mo ago
0