Token导航 LogoToken导航TokenDH.com
开发敏感数据clawhub未标认证来源可访问clear审计通过

chanjing-tts-voice-clone禅境 tts 语音克隆

Agent Skill

用于辅助音频、音乐、语音转写、语音合成或声音素材处理。它适合让 Agent 生成配乐说明、整理音频流程、调用语音工具或处理播客和视频配音素材。使用时需要确认输入音频来源、输出格式、时长和模型限制;涉及人声克隆、版权音乐或公开发布时,应先核对授权和合规边界。

总安装

6,512

周安装

266

GitHub Stars

公开资料未说明

下载量

2,107
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:chanjing-tts-voice-clone(禅境 tts 语音克隆)
来源仓库:https://github.com/iamzn1018/chanjing-tts-voice-clone
安装命令:
openclaw skills install chanjing-tts-voice-clone
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install chanjing-tts-voice-clone

简介

基于用户提供语音样本,利用禅境 TTS API 实现文本到语音的声音克隆。

  • 适用于播客配音、视频旁白与个性化语音合成需求。
  • 通过 clawhub 安装,需上传参考音频并指定文本内容进行合成。
  • 涉及人声克隆与版权素材使用时,应先核对授权边界与合规要求。
  • chanjing-tts-voice-clone 属于开发类 Skill,可作为该场景下的辅助能力补充。

SKILL.md

name
chanjing-tts-voice-clone
description
Use Chanjing TTS API to synthesize speech from text, using user-provided voice. Primary credential: credentials.json (app_id/secret_key; access_token and expire_in persisted on disk—do not commit; user accepts file-based secrets). Same credentials file as chanjing-credentials-guard. Not OpenClaw primaryEnv. Default path in metadata.openclaw.credentialModel. CHANJING_API_BASE and CHANJING_CONFIG_DIR optional. User supplies a public URL for reference audio (fetched by Chanjing servers). No ffmpeg/ffprobe required by this skill's scripts.
author
chan-skills
binaries
[]
env
category
媒体处理
tags
sibling_skills
credential_hint
~/.chanjing/credentials.json(可用 CHANJING_CONFIG_DIR 覆盖目录)
metadata
openclaw
homepage
https://doc.chanjing.cc
credentialModel
type
credentials_json
defaultPath
~/.chanjing/credentials.json
optionalEnv
apiBaseDefault
https://open-api.chanjing.cc
primaryEnvIntentionallyOmitted
true
persistAccessTokenOnDisk
true
sensitiveFields
doNotCommitToVcs
agentPolicy
alwaysSkill
false
modifiesOtherSkillsOrGlobalAgent
false

Chanjing TTS Voice Clone

技能包标识:chanjing-tts-voice-clone

功能说明

基于用户提供的参考音频公网 URL(由蝉镜服务端拉取)创建音色并合成语音;轮询任务并从接口返回 URL 下载结果。脚本依赖 ffmpeg/ffprobe。机器可读凭据声明见篇首 metadata.openclaw.credentialModelaccess_tokenexpire_in 会写入同一 credentials.json;刷新令牌时由本技能覆盖文件中对应字段,请勿将凭证提交到版本库。

运行依赖

  • python3 与同仓库 scripts/*.py
  • ffmpeg/ffprobe 门控

环境变量

  1. CHANJING_CONFIG_DIR(可选):credentials.json 所在目录,默认 ~/.chanjing
  2. CHANJING_API_BASE(可选):Open API 基址,默认 https://open-api.chanjing.cc

使用命令

  • ClawHub(slug 以注册表为准):clawhub run chanjing-tts-voice-clone
  • 本仓库python skills/chanjing-tts-voice-clone/scripts/create_voice.py …(见正文 How to Use

登记摘要(英文 · ClawHub / OpenClaw)

Primary credential (not primaryEnv): app_id + secret_key in ~/.chanjing/credentials.json (or $CHANJING_CONFIG_DIR/credentials.json); access_token and expire_in read/written in the same file (overwritten on token refresh). primaryEnv omitted by design. metadata.openclaw.credentialModel is the machine-readable registry source of truth.

Required vs optional: CHANJING_API_BASE optional (default https://open-api.chanjing.cc). CHANJING_CONFIG_DIR optional. No ffmpeg/ffprobe in skill metadata.

Purpose alignment: Voice clone TTSuser-supplied public URL for reference audio (Chanjing servers fetch it); then synthesize and download from API-returned URLs. Trust API host, response URLs, and user-chosen reference URL.

See How to UseReviewer Q&A (four items)Security & credentials (registry summary).

When to Use This Skill

Use this skill when the user needs to generate speech from text, with a user-provided reference voice. The reference audio needs to be provided as a publicly accessible url.

This TTS service supports:

  • bilingual Chinese and English
  • adjustment of speech rate
  • sentence-level timestamp

How to Use This Skill

前置条件(权限验证):执行本 Skill 前,必须先通过 chanjing-credentials-guard 完成 AK/SK 与 Token 校验。脚本与 guard 使用同一套凭证;无凭证时会执行 open_login_page.py 打开注册/登录页。

Reviewer Q&A (four items)

Direct answers for Purpose & Capability / Instruction Scope / Credentials / Persistence-style reviews; aligned with code, top Registry summary (English), and description frontmatter.

#TopicAnswer
1Purpose vs implementation; primary credentialAligned. Voice-clone TTS client; reference audio via user-provided public URL (server-side fetch). CHANJING_API_BASE not required. Primary credential in credentials.json; primaryEnv omitted. No ffmpeg/ffprobe in metadata.
2Runtime scope: secrets, URLsIn scope. Read/write credentials; browser if needed; HTTPS; user reference URL + output URLs from API—trust boundaries include user-chosen reference URL and API host.
3Env vs file secrets**CHANJING_* optional; app_id / secret_key / access_token / expire_in on disk—sensitive; token fields overwritten** on refresh.
4Persistence & privilegealways: false default; does not modify other skills or global agent config.

Security & credentials (registry summary)

Aligned with Reviewer Q&A above, top English registry summary, and description.

AspectDetails
Primary credentialSame file-based app_id / secret_key as other Chanjing skills. Not primaryEnv.
TokenRefreshed and written to ~/.chanjing/credentials.json (or CHANJING_CONFIG_DIR); access_token and expire_in are overwritten when the skill refreshes the token.
EnvOptional CHANJING_API_BASE, CHANJING_CONFIG_DIR.
User-supplied URLCreate Voice API takes a public URL to reference audio; the user is responsible for that URL’s safety and legality; Chanjing’s servers retrieve it.
Network / browserHTTPS to https://open-api.chanjing.cc; may open login when credentials are missing.
DownloadsGenerated audio is downloaded from URLs returned by the API—trust the API host.
PersistenceDefault always: false; does not modify other skills.

Chanjing-TTS-Voice-Clone provides an asynchronous speech synthesis API. Hostname for all APIs is: "https://open-api.chanjing.cc". All requests communicate using json. You should use utf-8 to encode and decode text throughout this task.

  1. Obtain an access\_token, which is required for subsequent requests
  2. Call the Create Voice API, which accepts a url to an audio file as reference voice
  3. Poll the Query Voice API until the result is success; keep the voice ID
  4. Call the Create Speech Generation Task API, using the voice ID, record the task_id
  5. Poll the Query Speech Generation Task Status API until the result is success
  6. When the task status is complete, use the corresponding url in API response to download the generated audio file

Get Access Token API

~/.chanjing/credentials.json 读取 app_idsecret_key,若无有效 Token 则调用:

POST /open/v1/access_token
Content-Type: application/json

请求体(使用本地配置的 app_id、secret_key):

{
  "app_id": "<从 credentials.json 读取>",
  "secret_key": "<从 credentials.json 读取>"
}

Response example:

{
  "trace_id": "8ff3fcd57b33566048ef28568c6cee96",
  "code": 0,
  "msg": "success",
  "data": {
    "access_token": "1208CuZcV1Vlzj8MxqbO0kd1Wcl4yxwoHl6pYIzvAGoP3DpwmCCa73zmgR5NCrNu",
    "expire_in": 1721289220
  }
}

Response field description:

First-level FieldSecond-level FieldDescription
codeResponse status code
msgResponse message
dataResponse data
access\_tokenValid for one day, previous token will be invalidated
expire\_inToken expiration time

Response status code description:

codeDescription
0Success
400Parameter format error
40000Parameter error
50000System internal error

Create Voice API

Post to the following endpoint to create a voice.

POST /open/v1/create_customised_audio
access_token: {{access_token}}
Content-Type: application/json

Request body example:

{
  "name": "example",
  "url": "https://example.com/abc.mp3"
}

Request field description:

FieldTypeRequiredDescription
namestringYesA name for this voice
urlstringYesurl to the reference audio file, format must be one of mp3, wav or m4a. Supported mime: audio/x-wav, audio/mpeg, audio/m4a, video/mp4. Size must not exceed 100MB. Recommended audio length: 30s-5min
model\_typestringYesUse "Cicada3.0-turbo"
languagestringNoEither "cn" or "en", default to "cn"

Response example:

{
  "trace_id": "2f0f50951d0bae0a3be3569097305424",
  "code": 0,
  "msg": "success",
  "data": "C-Audio-53e4e53ba1bc40de91ffaa74f20470fc"
}

Response field description:

FieldDescription
codeStatus Code
msgMessage
dataVoice ID, to be used in following steps

Response status code description:

CodeDescription
0Success
400Parameter Format Error
10400AccessToken Error
40000Parameter Error
40001QPS Exceeded
50000Internal System Error

Poll Voice API

Send a GET request to the following endpoint to query whether the voice is ready to be used, voice ID is obtained in the previous step. The polling process may take a few minutes, keep polling until the status indicates the voice is ready.

GET /open/v1/customised_audio?id={{voice_id}}
access_token: {{access_token}}

Response example:

{
  "trace_id": "7994cedae0f068d1e9e4f4abdf99215b",
  "code": 0,
  "msg": "success",
  "data": {
    "id": "C-Audio-53e4e53ba1bc40de91ffaa74f20470fc",
    "name": "声音克隆",
    "type": "cicada1.0",
    "progress": 0,
    "audio_path": "",
    "err_msg": "不支持的音频格式,请阅读接口文档",
    "status": 2
  }
}

Response field description:

First-level FieldSecond-level FieldDescription
codeStatus Code
msgResponse Message
data
idVoice ID
progressProgress: range 0-100
type
name
err\_msgError Message
audio\_path
status0-queued; 1-in progress; 2-done; 3-expired; 4-failed; 99-deleted

Response status code description:

CodeDescription
0Success
10400AccessToken Error
40000Parameter Error
40001QPS Exceeded
50000Internal System Error

Create Speech Generation Task API

Post to the following endpoint to submit a speech generation task:

POST /open/v1/create_audio_task
access_token: {{access_token}}
Content-Type: application/json

Request body example:

{
  "audio_man": "C-Audio-53e4e53ba1bc40de91ffaa74f20470fc",
  "speed": 1,
  "pitch": 1,
  "text": {
    "text": "Hello, I am your AI assistant."
  }
}

Request field description:

Parameter NameTypeNested KeyRequiredExampleDescription
audio\_manstringYesC-Audio-53e4e53ba1bc40de91ffaa74f20470fcVoice ID, obtained from previous step
speednumberYes1Speech rate, range: 0.5 (slow) to 2 (fast)
pitchnumberYes1Pitch (always set to 1)
textobjecttextYesHello, I am your Cicada digital humanRich text, text length limit less than 4,000 characters
aigc\_watermarkboolNofalseWhether to add visible watermark to audio, default is false

Response field description:

FieldDescription
codeResponse status code
msgResponse message
task\_idSpeech synthesis task ID

Example Response

{
  "trace_id": "dd09f123a25b43cf2119a2449daea6de",
  "code": 0,
  "msg": "success",
  "data": {
    "task_id": "88f635dd9b8e4a898abb9d4679e0edc8"
  }
}

Response status code description:

codeDescription
0Success
400Incoming parameter format error
10400AccessToken verification failed
40000Parameter error
40001Exceeds QPS limit
40002Production duration reached limit
50000System internal error

Query Speech Generation Task Status API

Post a request to the following endpoint:

POST /open/v1/audio_task_state
access_token: {{access_token}}
Content-Type: application/json

Request body example:

{
  "task_id": "88f635dd9b8e4a898abb9d4679e0edc8"
}

Request field description:

Parameter NameTypeRequiredExampleDescription
task\_idstringYes88f789dd9b8e4a121abb9d4679e0edc8Task ID obtained in previous step

Response body example:

{
  "trace_id": "ab18b14574bbcc31df864099d474080e",
  "code": 0,
  "msg": "success",
  "data": {
    "id": "9546a0fb1f0a4ae3b5c7489b77e4a94d",
    "type": "tts",
    "status": 9,
    "text": [
      "猫在跌落时能够在空中调整身体,通常能够四脚着地,这种”猫右自己“反射显示了它们惊人的身体协调能力和灵活性。核磁共振成像技术通过利用人体细胞中氢原子的磁性来生成详细的内部图像,为医学诊断提供了重要工具。"
    ],
    "full": {
      "url": "https://cy-cds-test-innovation.cds8.cn/chanjing/res/upload/tts/2025-04-08/093a59021d85a72d28a491f21820ece4.wav",
      "path": "093a59013d85a72d28a491f21820ece4.wav",
      "duration": 18.81
    },
    "slice": null,
    "errMsg": "",
    "errReason": "",
    "subtitles": [
      {
        "key": "20c53ff8cce9831a8d9c347263a400a54d72be15",
        "start_time": 0,
        "end_time": 2.77,
        "subtitle": "猫在跌落时能够在空中调整身体"
      },
      {
        "key": "e19f481b6cd2219225fa4ff67836448e054b2271",
        "start_time": 2.77,
        "end_time": 4.49,
        "subtitle": "通常能够四脚着地"
      },
      {
        "key": "140beae4046bd7a99fbe4706295c19aedfeeb843",
        "start_time": 4.49,
        "end_time": 5.73,
        "subtitle": "这种,猫右自己"
      },
      {
        "key": "e851881271876ab5a90f4be754fde2dc6b5498fd",
        "start_time": 5.73,
        "end_time": 7.97,
        "subtitle": "反射显示了它们惊人的身体"
      },
      {
        "key": "fbb0b4138bad189b9fc02669fe1f95116e9991b4",
        "start_time": 7.97,
        "end_time": 9.45,
        "subtitle": "协调能力和灵活性"
      },
      {
        "key": "f73404d135feaf84dd8fbea13af32eac847ac26d",
        "start_time": 9.45,
        "end_time": 12.49,
        "subtitle": "核磁共振成像技术通过利用人体"
      },
      {
        "key": "e18827931223962e477b14b2b8046947039ac222",
        "start_time": 12.49,
        "end_time": 14.77,
        "subtitle": "细胞中氢原子的磁性来生成"
      },
      {
        "key": "d137bf2b0c8b7a39e3f6753b7cf5d92bd877d2d9",
        "start_time": 14.77,
        "end_time": 15.97,
        "subtitle": "详细的内部图像"
      },
      {
        "key": "0773911ae0dbaa763a64352abdb6bdac3ff8f149",
        "start_time": 15.97,
        "end_time": 18.41,
        "subtitle": "为医学诊断提供了重要工具"
      }
    ]
  }
}

Response field description:

First-level FieldSecond-level FieldThird-level FieldDescription
codeResponse status code
msgResponse message
dataidAudio ID
typeSpeech type
statusStatus: 1 - in progress, 9 - done
textSpeech text
fullurlurl to download generated audio file
path
durationAudio duration
slice
errMsgError message
errReasonError reason
subtitles(array type)keySubtitle ID
start\_timeSubtitle start time point
end\_timeSubtitle end time point
subtitleSubtitle text

Response field description:

CodeDescription
0Response successful
10400AccessToken verification failed
40000Parameter error
50000System internal error

Scripts

本 Skill 提供脚本(skills/chanjing-tts-voice-clone/scripts/),与 chanjing-credentials-guard 使用同一配置文件;无 AK/SK 时会执行 open_login_page.py 脚本打开注册/登录页。

脚本说明
create_voice.py提交定制声音任务(参考音频 URL),输出 voice_id
poll_voice.py轮询定制声音直到就绪(status=2),输出 voice_id
create_task.py使用定制声音创建 TTS 任务,输出 task_id
poll_task.py轮询 TTS 任务直到完成,输出音频下载 URL

示例(在项目根或 skill 目录下执行):

# 1. 创建定制声音(参考音频需为公开 URL)
VOICE_ID=$(python skills/chanjing-tts-voice-clone/scripts/create_voice.py --name "我的声音" --url "https://example.com/ref.mp3")

# 2. 轮询直到声音就绪
python skills/chanjing-tts-voice-clone/scripts/poll_voice.py --voice-id "$VOICE_ID"

# 3. 创建 TTS 任务
TASK_ID=$(python skills/chanjing-tts-voice-clone/scripts/create_task.py --audio-man "$VOICE_ID" --text "Hello, I am your AI assistant.")

# 4. 轮询到完成,得到音频下载链接
python skills/chanjing-tts-voice-clone/scripts/poll_task.py --task-id "$TASK_ID"

适合场景

01

OpenClaw 用户查找和安装 Skill 时

02

用户想查找某类 Agent Skill 时

03

需要根据任务场景推荐可安装能力包时

04

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

补充不同宿主或平台的使用分布数据

能力 5

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

86.05%
按下载量换算1,813

安全审计

VirusTotal

通过

ClawScan

通过

Static analysis

通过

权限和风险

敏感数据

该 Skill 可能接触密钥、Token、环境变量或敏感配置,应进入高风险复核队列,默认不自动发布。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills