15 skills found
aaronnat23 / voice-opsSelf-hosted AI workspace where chat becomes visual workflows, multi-agent operations, and reviewable automations. Local memory; local or cloud models
huangjunsen0406 / py-xiaozhiOpen-source AI assistant ecosystem with MCP integrations, multimodal workflows, IoT support, and cross-platform voice interaction.
NPC-Worldwide / npcpyThe python library for research and development in NLP, multimodal LLMs, Agents, ML, Knowledge Graphs, and more.
NarratorAI-Studio / narrator-ai-cli-skillAI 解说大师 — Agent skill;封装 narrator-ai-cli 供 Claude/Codex 等工具调用
mbailey / voicemodeNatural voice conversations with Claude Code
Remy2404 / PolymindA powerful, multi-modal Telegram bot leveraging cutting-edge AI technologies including Gemini, DeepSeek, OpenRouter, and 50+ AI models for comprehensive conversational assistance, media processing, and collaborative features with MCP (Model Context Protocol) integration.
waxberry-dev / live-translate-mcpMCP server for local speech translation (EN ↔ 中文) via Whisper + Claude + Piper
majiayu000 / agent-sona-learning-optimizerAgent skill for sona-learning-optimizer - invoke with $agent-sona-learning-optimizer
Intradyne / AnyLoom-AnythingLLM-Local-AI-agentic-DyTopo-swarmChatGPT-like AI that runs 100% locally on your hardware. No subscriptions, no cloud, complete privacy. Multi-agent swarm + 10 MCP tools + hybrid RAG vector DB + . Runs on one GPU (RTX 5090 recommended)
arukaraz / synthseekSelf-hosted music discovery library manager with modern web UI, MCP tools and more...
howdoiusekeyboard / TrueVoice-MCPA Model Context Protocol server that helps AI generate human-like text without AI slop
kyegomez / OpenMythosA theoretical reconstruction of the Claude Mythos architecture, built from first principles using the available research literature.
second-state / audio-ttsGenerate speech audio from text using Qwen3 TTS, or clone a voice from reference audio. Triggered when the user wants to convert text to speech, generate audio, read text aloud, or clone/mimic a voice. Supports multiple speakers, English and Chinese, and emotion/style control.
AI272 / speakerSpeaker is a Codex skill project for academic presentations: read real.pptx, combine text extraction, PPTX structure parsing, page-by-page rendering, OCR, and visual review to generate page-by-page speaker notes, and write a clean version of the lecture into the PowerPoint comment area.
Ryanzucchi / analise-consistencia-narrativa-tempo-e-eventosGerencia e valida a consistencia causal de tramas ramificadas e linhas do tempo alternativas usando grafos de eventos.