speech-to-text
Transcribe video to timestamped text using Whisper tiny model (pre-installed).
Install / Use
npx skills add benchflow-ai/skillsbench --skill speech-to-textInstalls into whichever agent you are using.
SKILL.md
Installable skill definition
Quality Score
Category
Development & EngineeringSupported Platforms
Our assessment of speech-to-text
speech-to-text scores 57/100 on our quality scale, 3542nd of 3,997 Development & Engineering skills we index.
Its SKILL.md is 487 bytes long, lightly structured (2 headings) with 2 code examples: very short, closer to a stub than a full skill.
With 1,813 GitHub stars, it is one of the more widely adopted skills in the catalogue.
Maintenance, license and trust
- The repository was last updated about 2 months ago, so speech-to-text is actively maintained.
- It is released under the Apache-2.0 license, a permissive license that allows use, modification and commercial use with attribution.
- Its trust signals score 100/100, with no cautions. These come from repository metadata, not a code audit — read the skill file before letting an agent act on it.
speech-to-text compared with similar skills
All 4 of these similar skills score higher than speech-to-text; compare them before choosing.
| Skill | Score | Stars | Updated | Format |
|---|---|---|---|---|
| speech-to-text (this skill)by benchflow-ai | 57 | 1.8k | 2mo ago | SKILL.md |
| Agent-Reachby Panniantong | 100 | 86.4k | 15d ago | CLAUDE.md |
| headroomby headroomlabs-ai | 100 | 74.2k | today | CLAUDE.md |
| ai-job-searchby MadsLorentzen | 100 | 44.6k | today | CLAUDE.md |
| claude-howtoby luongnv89 | 100 | 41.7k | today | CLAUDE.md |
Frequently asked questions
- How do I install speech-to-text?
- Run
npx skills add benchflow-ai/skillsbench --skill speech-to-text. The install tabs above show the steps for each supported agent. - Which AI agents does speech-to-text work with?
- It is written for Universal, as a SKILL.md file. Other agents that read the same format can often use it too.
- Is speech-to-text safe to use?
- It is Apache-2.0-licensed and scores 100/100 on trust signals. Skills are instructions an agent will follow, so read the file before installing it and do not approve commands you do not understand.
- Is speech-to-text still maintained?
- The repository was last updated about 2 months ago, so speech-to-text is actively maintained.
Skill content
View source on GitHubname: speech-to-text description: Transcribe video to timestamped text using Whisper tiny model (pre-installed).
Speech-to-Text
Transcribe video to text with timestamps.
Usage
python3 scripts/transcribe.py /root/tutorial_video.mp4 -o transcript.txt --model tiny
This produces output like:
[0.0s - 5.2s] Welcome to this tutorial.
[5.2s - 12.8s] Today we're going to learn...
The tiny model is pre-downloaded and takes ~2 minutes for a 23-min video.
Related Skills
Agent-Reach
86.4kGive your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili, XiaoHongShu — one CLI, zero API fees.
headroom
74.2kCompress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
ai-job-search
44.6kThe job search that runs on your machine. AI job application framework built on Claude Code: evaluate postings, tailor CVs, write cover letters, prep interviews. Fork it and own it.
claude-howto
41.7kA visual, example-driven guide to Claude Code — from basic concepts to advanced agents, with copy-paste templates that bring immediate value.
Languages
Trust signals
From repository metadata: license, adoption, age and documentation. Not a code audit — see the Safety scan above for what the skill file itself contains.
