12 skills found
modelscope / FunASROpen-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
howdoiusekeyboard / TrueVoice-MCPA Model Context Protocol server that helps AI generate human-like text without AI slop
alessandro9110 / Speech-To-Text-With-DatabricksAn end-to-end, scalable STT solution on Databricks that transcribes audio into structured text in Delta Lake, ready for analytics, search, and GenAI/RAG.
aws-samples / recorded-voice-insight-extraction-webappA generative AI tool to boost productivity by transcribing and analyzing audio or video recordings containing speech
waxberry-dev / live-translate-mcpMCP server for local speech translation (EN ↔ 中文) via Whisper + Claude + Piper
second-state / audio-ttsGenerate speech audio from text using Qwen3 TTS, or clone a voice from reference audio. Triggered when the user wants to convert text to speech, generate audio, read text aloud, or clone/mimic a voice. Supports multiple speakers, English and Chinese, and emotion/style control.
SEACrowd / nova-sonicEvaluating Conversational Agents in a Multimodal Multilingual Environment
shreyas-s-rao / claude-code-narratorA text-to-speech plugin for Claude Code.
cafferychen777 / ChatSpatialMCP server for spatial transcriptomics analysis through natural language interfaces.
huangjunsen0406 / py-xiaozhiOpen-source AI assistant ecosystem with MCP integrations, multimodal workflows, IoT support, and cross-platform voice interaction.
78 / xiaozhi-esp32An MCP-based chatbot | 一个基于MCP的聊天机器人
WPeace-HcH / wpegpt-analyzer驱动 IDA 配合 WPeGPT 插件对二进制可执行文件(PE/ELF)进行自动化逆向分析,输出包含程序用途、网络 IoC、可疑函数及漏洞评估的结构化报告。