15 skills found
KuaaMU / mcp-vision-bridgeMCP server that gives text-only LLM coding agents vision — analyze images via any multimodal model (mimo, Claude, Gemini, OpenAI-compatible). Works with Claude Code, Codex, Kimi, opencode, PI.
ProjectLiminality / DreamTalkA programmatic animation library extending the ancient modality of SandTalk into the digital domain
FishWoWater / trellis_mcpModel Context Protocol(MCP) for TRELLIS(SOTA text-to-3d/image-to-3d) models
SamurAIGPT / Generative-Media-SkillsMulti-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image, video, and audio generation powered by muapi.ai.
kyegomez / OpenMythosA theoretical reconstruction of the Claude Mythos architecture, built from first principles using the available research literature.
semaj90 / 05-output-boundarylegal ai contextual chat
olimorris / codecompanion.nvim✨ AI Coding, Vim Style
symgraph / BinAssistMCPBinary Ninja plugin to provide MCP functionality.
Agents365-ai / mermaid-skillMermaid diagrams (.mmd) from natural language with validation loop. 11+ types, multi-backend (mmdc / Kroki), PNG/SVG/PDF, multi-agent.
SEACrowd / nova-sonicEvaluating Conversational Agents in a Multimodal Multilingual Environment
bytedance / UI-TARS-desktopThe Open-Source Multimodal AI Agent Stack: Connecting Cutting-Edge AI Models and Agent Infra
mcpland / storybook-mcpA MCP server for Storybook.
michaeltrilford / muiscan-mcpTranslate Figma files to MUI Design System via the Muiscan MCP Server
miantiao-me / bm.md更好用的 Markdown 排版助手|一键适配微信公众号、网页与图片。
kitlau86 / agent-vision-mcpAn MCP server that gives non-vision LLMs the ability to "see" images. Plug in any OpenAI-compatible vision API (Gemini, Qwen-VL, OpenAI, or self-hosted) and your text-only model — like DeepSeek in Claude Code — can analyze screenshots, OCR text, read charts, and more.