40 skills found · Page 2 of 2
Anil-matcha / Veo-4-APIPython wrapper for Veo 4 API by Google DeepMind — native 4K AI video with integrated audio, character consistency & advanced camera controls.
arizawan / vidlizerPoint it at a video, image, or PDF — get structured JSON. uvx vidlizer[mcp]. Runs local (Ollama/gemma4, LM Studio, oMLX) or cloud (OpenRouter). CLI + MCP server for Claude Code, Cursor, and Claude Desktop.
mohsinkhadim59 / youtube-obsidian-mcpAutomatically generate professional Obsidian study notes from YouTube videos using Claude AI. Extract subtitles, capture screenshots at key timestamps, and create beautifully formatted markdown notes. Perfect for learning programming, design, academics, business - any topic on YouTube!
HarperZ9 / gatherResearch intake that reaches the hard places: web, video, papers, scanned PDFs, browser, OCR, and audio into structured research packets. DOM extraction and change tracking built in; provenance rides along on every item.
compozy / kbComprehensive skill for the `kb` CLI and the Karpathy Knowledge Base pattern. Covers the full KB lifecycle — topic scaffolding, multi-source ingestion (URLs, files, YouTube videos and channels, Instagram reels, bookmarks, codebases), wiki article compilation, cross-article querying with file-back, l…
aicw-io / aicw-videoAICW Video is made for enhancing videos with humans by detecting key moments, adding captions and voice overs, adding side illustrations based on what is mentioned in video
Rimagination / dy-noteCodex Skill for turning Douyin videos into evidence-based Markdown notes
yizhiyanhua-ai / youtube-ai-digestClaude Code Skill: Browse, summarize and capture AI-related YouTube videos
mocchalera / compile-timelineRoughCut Agent — Autonomous video editing agent that runs intent → analysis → triage → blueprint → compile → review in one shot
sakamoto-family-smile / videodbSee, Understand, Act on video and audio. See- ingest from local files, URLs, RTSP/live feeds, or live record desktop; return realtime context and playable stream links. Understand- extract frames, build visual/semantic/temporal indexes, and search moments with timestamps and auto-clips.