21 skills found
mudler / LocalAILocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
heygen-com / hyperframesWrite HTML. Render video. Built for agents.
palmier-io / palmier-promacOS video editor built for AI
netease-youdao / LobsterAIOpen-source, desktop-grade AI agent that gets real work done — data analysis, slides, docs, video & web research. Built on OpenClaw; runs tools on your real desktop and takes commands from your phone via WeChat, Feishu, DingTalk & Telegram.
NVIDIA-AI-Blueprints / video-search-and-summarizationNVIDIA AI Blueprint for video search and summarization (VSS) is a GPU-accelerated reference architecture for building video analytics agents with real-time verified alerts, visual Q&A, and automated reporting.
MiniMax-AI / MiniMax-MCPOfficial MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.
kimsungwhee / apple-docs-mcpMCP server for Apple Developer Documentation - Search iOS/macOS/SwiftUI/UIKit docs, WWDC videos, Swift/Objective-C APIs & code examples in Claude, Cursor & AI assistants
gyoridavid / short-video-makerCreates short videos for TikTok, Instagram Reels, and YouTube Shorts using the Model Context Protocol (MCP) and a REST API.
jordanrendric / claude-video-visionGive Claude the ability to watch and understand videos — Claude Code plugin with frame extraction and multimodal audio analysis
artokun / comfyui-mcpLocal-first, agent-native control plane for ComfyUI — MCP server + sidebar agent that generates images, video & audio, authors and runs workflows, and edits your live graph in natural language on ANY LLM (Claude, ChatGPT, Gemini, offline Ollama, or any hosted model).
TeleAI-UAGI / telememTeleMem is a high-performance drop-in replacement for Mem0, featuring semantic deduplication, long-term dialogue memory, and multimodal video reasoning.
hetpatel-11 / Adobe_Premiere_Pro_MCPAdobe Premiere Pro MCP. 282 tools for AI-driven video editing via MCP, for Codex, Claude, and other MCP clients.
PixVerseAI / PixVerse-MCPOfficial PixVerse Model Context Protocol (MCP) server that enables interaction with powerful AI video generation APIs.
merterbak / Grok-MCPMCP server for xAI’s Grok API with Web/X search, vision, image/video generation and file support
guimatheus92 / mcp-video-analyzerMCP server that turns any video — YouTube, Instagram, TikTok, Loom, X, Vimeo, direct URLs, local files — into transcripts, key frames, OCR text, and metadata for AI agents.
totigm / humanjsHumanize browser automation for AI agents, QA tests, and demos. Playwright-first — record sessions to video or runnable tests, and drive it over MCP from your favorite AI.
Anil-matcha / Veo-4-APIPython wrapper for Veo 4 API by Google DeepMind — native 4K AI video with integrated audio, character consistency & advanced camera controls.
arizawan / vidlizerPoint it at a video, image, or PDF — get structured JSON. uvx vidlizer[mcp]. Runs local (Ollama/gemma4, LM Studio, oMLX) or cloud (OpenRouter). CLI + MCP server for Claude Code, Cursor, and Claude Desktop.
mohsinkhadim59 / youtube-obsidian-mcpAutomatically generate professional Obsidian study notes from YouTube videos using Claude AI. Extract subtitles, capture screenshots at key timestamps, and create beautifully formatted markdown notes. Perfect for learning programming, design, academics, business - any topic on YouTube!
HarperZ9 / gatherResearch intake that reaches the hard places: web, video, papers, scanned PDFs, browser, OCR, and audio into structured research packets. DOM extraction and change tracking built in; provenance rides along on every item.
aicw-io / aicw-videoAICW Video is made for enhancing videos with humans by detecting key moments, adding captions and voice overs, adding side illustrations based on what is mentioned in video