音频处理低风险

OpenAI Whisper 语音转文字

通过本地 Whisper CLI 进行离线语音转文字,无需 API Key。

by @openaiv1.0.06 浏览3 下载
#audio#cli#local#openai#speech-to-text#subtitle#transcription
推荐方式

直接交给 AI

无需安装文件。复制一句话,让 AI 在线读取这个 Skill。

下载 Skill ZIP ↓ 查看 SKILL.md 原文 ↗
01

它能帮你做什么

先看懂,再决定要不要交给 AI。

本 Skill 封装了 OpenAI Whisper 命令行工具的使用方式,允许在本地对音频文件进行语音识别和翻译。支持多种输出格式(txt、srt 等),模型首次运行时会自动下载到 ~/.cache/whisper 缓存目录。可选择不同规模的模型以在速度与准确度之间权衡。

✓将录音音频转写为文本
✓生成视频字幕(SRT)
✓对英文音频进行翻译转写
✓在不联网或不使用 API 的情况下完成语音转文字
02

怎么交给 AI

在线读取优先,本地安装作为备选。

◎
在线读取推荐 · 不需要安装

适合能访问网页的 ChatGPT、Agent 或其他 AI。

AI Prompt请访问 https://skills.dhmip.cn/skills/openai/openai-whisper/SKILL.md,读取并按照该 Skill 完成任务;如当前环境支持本地安装,也可以下载该 Skill。
↓
下载安装到 Agent适合支持 Skills 的客户端

未登录时可使用公共安装文档;登录后可以按不同 AI 分开管理。

Install Prompt请根据 https://skills.dhmip.cn/install/skillhub.md,安装 @openai/openai-whisper。
登录后管理多个 AI →
03

兼容性与要求

安装或使用前,先确认环境是否匹配。

适用客户端

clawdbot

使用要求

  • 系统已安装 whisper CLI(可通过 brew install openai-whisper 安装)
  • 首次运行需联网下载模型到 ~/.cache/whisper
  • 足够的磁盘空间存放所选模型
  • 足够的内存/CPU 资源以运行所选规模的模型
⌘技术详情查看完整 SKILL.md 与原始内容
⌄
SKILL.mdRaw ↗

name: openai-whisper
description: Local speech-to-text with the Whisper CLI (no API key).
homepage: https://openai.com/research/whisper
metadata: {"clawdbot":{"emoji":"🎙️","requires":{"bins":["whisper"]},"install":[{"id":"brew","kind":"brew","formula":"openai-whisper","bins":["whisper"],"label":"Install OpenAI Whisper (brew)"}]}}


Whisper (CLI)

Use whisper to transcribe audio locally.

Quick start

  • whisper /path/audio.mp3 --model medium --output_format txt --output_dir .
  • whisper /path/audio.m4a --task translate --output_format srt

Notes

  • Models download to ~/.cache/whisper on first run.
  • --model defaults to turbo on this install.
  • Use smaller models for speed, larger for accuracy.