音频处理低风险
OpenAI Whisper 语音转文字
通过本地 Whisper CLI 进行离线语音转文字,无需 API Key。
01
它能帮你做什么
先看懂,再决定要不要交给 AI。
本 Skill 封装了 OpenAI Whisper 命令行工具的使用方式,允许在本地对音频文件进行语音识别和翻译。支持多种输出格式(txt、srt 等),模型首次运行时会自动下载到 ~/.cache/whisper 缓存目录。可选择不同规模的模型以在速度与准确度之间权衡。
✓将录音音频转写为文本
✓生成视频字幕(SRT)
✓对英文音频进行翻译转写
✓在不联网或不使用 API 的情况下完成语音转文字
02
怎么交给 AI
在线读取优先,本地安装作为备选。
在线读取推荐 · 不需要安装
适合能访问网页的 ChatGPT、Agent 或其他 AI。
AI Prompt
请访问 https://skills.dhmip.cn/skills/openai/openai-whisper/SKILL.md,读取并按照该 Skill 完成任务;如当前环境支持本地安装,也可以下载该 Skill。下载安装到 Agent适合支持 Skills 的客户端
未登录时可使用公共安装文档;登录后可以按不同 AI 分开管理。
Install Prompt
登录后管理多个 AI →
请根据 https://skills.dhmip.cn/install/skillhub.md,安装 @openai/openai-whisper。03
兼容性与要求
安装或使用前,先确认环境是否匹配。
适用客户端
clawdbot
使用要求
- 系统已安装 whisper CLI(可通过 brew install openai-whisper 安装)
- 首次运行需联网下载模型到 ~/.cache/whisper
- 足够的磁盘空间存放所选模型
- 足够的内存/CPU 资源以运行所选规模的模型
⌘技术详情查看完整 SKILL.md 与原始内容⌄
name: openai-whisper
description: Local speech-to-text with the Whisper CLI (no API key).
homepage: https://openai.com/research/whisper
metadata: {"clawdbot":{"emoji":"🎙️","requires":{"bins":["whisper"]},"install":[{"id":"brew","kind":"brew","formula":"openai-whisper","bins":["whisper"],"label":"Install OpenAI Whisper (brew)"}]}}
Whisper (CLI)
Use whisper to transcribe audio locally.
Quick start
whisper /path/audio.mp3 --model medium --output_format txt --output_dir .whisper /path/audio.m4a --task translate --output_format srt
Notes
- Models download to
~/.cache/whisperon first run. --modeldefaults toturboon this install.- Use smaller models for speed, larger for accuracy.