Token导航 LogoToken导航TokenDH.com
效率执行命令clawhub未标认证来源可访问clear审计通过

openclaw-whisper-voiceOpenClaw Whisper voice 音频

Agent Skill

用于辅助音频、音乐、语音转写、语音合成或声音素材处理。它适合让 Agent 生成配乐说明、整理音频流程、调用语音工具或处理播客和视频配音素材。使用时需要确认输入音频来源、输出格式、时长和模型限制;涉及人声克隆、版权音乐或公开发布时,应先核对授权和合规边界。

总安装

5,784

周安装

241

GitHub Stars

公开资料未说明

下载量

1,928
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:openclaw-whisper-voice(OpenClaw Whisper voice 音频)
来源仓库:https://github.com/sabyaghosh/openclaw-whisper-voice
安装命令:
openclaw skills install openclaw-whisper-voice
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install openclaw-whisper-voice

简介

提供 OpenClaw 网关上的本地 Whisper 语音转文本。

  • 适用于 WhatsApp、Telegram 等平台的入站语音注释处理。
  • 支持音频文件转录和实时转换,提升语音处理能力。
  • 安装命令:openclaw skills install openclaw-whisper-voice。
  • 需确认输入音频来源和输出格式,避免版权或隐私风险。

SKILL.md

name
openclaw-whisper-voice
description
Local Whisper speech-to-text for audio files and inbound voice notes on the OpenClaw Gateway host. Use when setting up local transcription for WhatsApp, Telegram, or other audio attachments; when configuring tools.media.audio with a CLI fallback instead of a cloud API; or when you need a reusable shell entrypoint that makes Whisper + ffmpeg work reliably on Linux.
metadata

OpenClaw Whisper Voice

Use this skill to make local Whisper transcription dependable on the OpenClaw Gateway host.

Install on the host

Run:

{baseDir}/scripts/install_local_whisper.sh

The installer:

  • installs Python packages into ~/.local
  • installs a CPU-safe PyTorch build
  • installs openai-whisper
  • installs imageio-ffmpeg
  • creates stable ~/.local/bin/whisper and ~/.local/bin/ffmpeg launchers

Transcribe a file manually

Use the wrapper instead of raw whisper when reliability matters:

{baseDir}/scripts/transcribe.sh /path/to/audio.ogg
{baseDir}/scripts/transcribe.sh /path/to/audio.m4a --model tiny --stdout-only
{baseDir}/scripts/transcribe.sh /path/to/audio.mp3 --task translate --format srt

Configure inbound WhatsApp and Telegram voice notes

Patch OpenClaw config so inbound audio uses the wrapper:

{
  tools: {
    media: {
      audio: {
        enabled: true,
        maxBytes: 20971520,
        timeoutSeconds: 120,
        models: [
          {
            type: "cli",
            command: "{baseDir}/scripts/transcribe.sh",
            args: ["{{MediaPath}}", "--model", "base", "--stdout-only"],
            timeoutSeconds: 120
          }
        ]
      }
    }
  }
}

Model choices

  • tiny: fastest, weakest accuracy
  • base: best default for chat voice notes
  • small or larger: better accuracy, heavier CPU and RAM use

Output rules

  • Use --stdout-only for tools.media.audio so stdout is only transcript text.
  • Use --format txt|srt|vtt|json for standalone file transcription.
  • First model download goes into ~/.cache/whisper.

适合场景

01

OpenClaw 用户查找和安装 Skill 时

02

用户想查找某类 Agent Skill 时

03

需要根据任务场景推荐可安装能力包时

04

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

补充不同宿主或平台的使用分布数据

能力 5

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

86.03%
按下载量换算1,659

安全审计

VirusTotal

通过

ClawScan

通过

Static analysis

通过

权限和风险

执行命令

安装流程涉及命令执行,可能通过 openclaw skills install openclaw-whisper-voice 联网下载 Skill 或依赖。用户安装前应确认命令来源、仓库内容和执行环境。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills