15 skills found
sailscastshq / pelliculeCreate videos with Vue using Pellicule - a Vue-native video rendering library
AgriciDaniel / claude-videoAI-powered video production suite for Claude Code. Edit, transcode, caption, analyze, generate (Veo), stock footage promos with contrast-aware text, shortform pipeline, and more.
waxberry-dev / live-translate-mcpMCP server for local speech translation (EN ↔ 中文) via Whisper + Claude + Piper
Agents365-ai / yt2bbYouTube to Bilibili video repurposing with bilingual subtitles
sakamoto-family-smile / videodbSee, Understand, Act on video and audio. See- ingest from local files, URLs, RTSP/live feeds, or live record desktop; return realtime context and playable stream links. Understand- extract frames, build visual/semantic/temporal indexes, and search moments with timestamps and auto-clips.
Vincentwei1021 / video-shotcraftAI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template
jordanrendric / claude-video-visionGive Claude the ability to watch and understand videos — Claude Code plugin with frame extraction and multimodal audio analysis
JimmySadek / youtube-fetcher-to-markdownClaude Code skill: turn YouTube videos into structured, Obsidian-ready Markdown notes with full metadata, chapters, and transcripts
alessandro9110 / Speech-To-Text-With-DatabricksAn end-to-end, scalable STT solution on Databricks that transcribes audio into structured text in Delta Lake, ready for analytics, search, and GenAI/RAG.
AgriciDaniel / claude-shortsInteractive longform-to-shortform video creator — Claude Code skill with Remotion-rendered animated captions, AI segment scoring, cursor tracking, and audio-aware boundary snapping
programmerloverun / bilibili-to-doc🎬 哔哩哔哩视频转文档 | Claude Code Skill: 自动将B站视频AI字幕提取为结构化Markdown文档
second-state / audio-ttsGenerate speech audio from text using Qwen3 TTS, or clone a voice from reference audio. Triggered when the user wants to convert text to speech, generate audio, read text aloud, or clone/mimic a voice. Supports multiple speakers, English and Chinese, and emotion/style control.
shehryarsaroya / agenttransferOpen-source infrastructure for AI agents — each gets an email address, folder, inbox, and API key. Move files through named inboxes instead of shared cloud credentials; signed receipts and MCP built in.
guimatheus92 / mcp-video-analyzerMCP server that turns any video — YouTube, Instagram, TikTok, Loom, X, Vimeo, direct URLs, local files — into transcripts, key frames, OCR text, and metadata for AI agents.
beeswaxpat / ffmpeg-render-proParallel video rendering with live dashboard, GPU auto-detection, and stream-copy concat. MCP server with 7 typed tools for AI agents, a Claude Code skill, and a CLI.