17 skills found
YouMind-OpenLab / nano-banana-pro-prompts-recommend-skillAI skill for OpenClaw & Claude Code — recommend from 10000+ Nano Banana Pro (Gemini) image prompts. Smart search by use case, content remix, sample images.
artokun / comfyui-mcpLocal-first, agent-native control plane for ComfyUI — MCP server + sidebar agent that generates images, video & audio, authors and runs workflows, and edits your live graph in natural language on ANY LLM (Claude, ChatGPT, Gemini, offline Ollama, or any hosted model).
Bria-AI / bria-skillClaude Code skills for Bria AI - generate, edit, and transform images with Fibo, RMBG-2.0, and VGL structured prompts
SpillwaveSolutions / plantumlA Plantuml Claude Skill that can generate images and help you create Plantuml digrams from source code. It can also extract plantuml diagrams from a markdown file and then generate each of those to images and create a new markdown file with image links to those diagrams.
cbtw-apac / qdrant-loaderEnterprise-ready vector database toolkit for building searchable knowledge bases from multiple data sources. Supports multi-project management, automatic ingestion from Confluence/JIRA/Git, intelligent file conversion (PDF/Office/images), and semantic search.
lalanikarim / comfy-mcp-serverA server using FastMCP framework to generate images based on prompts via a remote Comfy server.
Shaohan-He / deepseek-eyes给 DeepSeek 装上眼睛 — MCP Server + 通义千问VL, 剪贴板图片→视觉模型→文字描述 / Give DeepSeek the ability to see images via clipboard + Qwen-VL
albertzhangz10 / design-system-skillClaude Code skill: Generate design.md, design-guidelines.md & design-components.md from any design references — images, PDFs, links, screenshots. For AI-assisted coding (Cursor, Claude Code, Copilot)
wells1137 / media-skillsA collection of open-source Agent Skills for content creation — images, audio, and video.
kitlau86 / agent-vision-mcpAn MCP server that gives non-vision LLMs the ability to "see" images. Plug in any OpenAI-compatible vision API (Gemini, Qwen-VL, OpenAI, or self-hosted) and your text-only model — like DeepSeek in Claude Code — can analyze screenshots, OCR text, read charts, and more.
dazeb / wikipedia-mcp-image-crawlerA Wikipedia Image Search Tool. Follows Creative Commons Licences for images and uses them in your projects via Claude Desktop/Cline.
KuaaMU / mcp-vision-bridgeMCP server that gives text-only LLM coding agents vision — analyze images via any multimodal model (mimo, Claude, Gemini, OpenAI-compatible). Works with Claude Code, Codex, Kimi, opencode, PI.
kbarbel640-del / nvidia-image-genGenerate and edit images using NVIDIA FLUX models
Trompetilla / nvidia-image-genGenerate and edit images using NVIDIA FLUX models
johnalbertini14-glitch / nvidia-image-genGenerate and edit images using NVIDIA FLUX models
mits-pl / seo-imagesOpen-source AI coding agent. Desktop app, bring your own model. Writes code, browses the web, verifies its work. Apache 2.0.
yantianqi1 / comic-studio-workflowUse when the user wants to create a comic/manga from a story, premise, or prompt. Executes the full pipeline: story analysis, character design, storyboard planning, image prompt composition, and image generation. Produces final comic page images as output.