ElevenLabs · 音訊與音樂

Hermes Agent x ElevenLabs

ElevenLabs 是語音 AI 公司,其模型涵蓋 TTS、對話、音效、轉錄及音訊分離。透過 RunAPI,所有 ElevenLabs 端點共用同一個 key,並按次計費。

6 variants 由 $0.04 / 分鐘 可商用

Prerequisite: npx runapi mcp install

Prompt

Prompt 模型

任務合適時使用 audio-isolation:從混合音訊來源中提取人聲。

You have access to RunAPI task tools for ElevenLabs.

Available ElevenLabs models:
- audio-isolation: 從混合音訊來源中提取人聲
  endpoints: /api/v1/elevenlabs/isolate_audio
  request fields: source_audio_url
- sound-effect-v2: 文字生成音效,適用於遊戲、影片及 podcast
  endpoints: /api/v1/elevenlabs/text_to_sound
  request fields: text, loop, duration_seconds, prompt_influence, output_format
- speech-to-text: 支援 29 種以上語言的轉錄,具說話者分辨功能
  endpoints: /api/v1/elevenlabs/speech_to_text
  request fields: source_audio_url, language_code, diarize
- text-to-dialogue-v3: 多人對話生成,自然輪流說話
  endpoints: /api/v1/elevenlabs/text_to_dialogue
  request fields: dialogue, stability, language_code
- text-to-speech-multilingual-v2: 29 種語言;情感表達最逼真;適合有聲書
  endpoints: /api/v1/elevenlabs/text_to_speech
  request fields: model, text, voice, language_code
- text-to-speech-turbo-v2.5: 32 種語言;非英語快 3 倍;40K 字元上限
  endpoints: /api/v1/elevenlabs/text_to_speech
  request fields: model, text, voice, language_code

Use the model ID and endpoint that match the user's request.
Example prompt: 將這段文字轉換為自然語音,使用溫暖的英式男聲,語速適中。
/api/v1/elevenlabs/isolate_audio: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/isolate_audio,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/isolate_audio/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/elevenlabs/text_to_sound: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/text_to_sound,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/text_to_sound/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/elevenlabs/speech_to_text: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/speech_to_text,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/speech_to_text/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/elevenlabs/text_to_dialogue: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/text_to_dialogue,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/text_to_dialogue/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/elevenlabs/text_to_speech: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/text_to_speech,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/text_to_speech/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。

公開版本與端點

模型 ID 端點 起始價格 模型目錄
audio-isolation
/api/v1/elevenlabs/isolate_audio
$0.12 / 分鐘 模型詳情
sound-effect-v2
/api/v1/elevenlabs/text_to_sound
$0.15 / 分鐘 模型詳情
speech-to-text
/api/v1/elevenlabs/speech_to_text
$0.04 / 分鐘 模型詳情
text-to-dialogue-v3
/api/v1/elevenlabs/text_to_dialogue
$0.14 / 1K 字元 模型詳情
text-to-speech-multilingual-v2
/api/v1/elevenlabs/text_to_speech
$0.12 / 1K 字元 模型詳情
text-to-speech-turbo-v2.5
/api/v1/elevenlabs/text_to_speech
$0.06 / 1K 字元 模型詳情

驗證

輪詢,直到 Task 進入終態

選擇 <model-id> 後產生驗證命令。

設定

指南端點: <endpoint>

選擇 <model-id> 後,依端點的公開輸入契約產生請求。
使用流程

三步開始使用

  1. 選擇模型 ID

    選擇公開模型 ID,並查看其端點與目前起始價格。

  2. 設定 RunAPI

    呼叫端點前設定 RUNAPI_API_KEY。

  3. 驗證結果

    對非同步端點輪詢同一路徑,直到 Task 進入終態。

使用 Hermes Agent + ElevenLabs 可以建立甚麼

  • 對話式語音 Agent

    建立說話自然的語音 Agent,以低延遲語音支援客服機器人、助理或電話界面。

  • YouTube 內容旁白

    為 YouTube 影片製作旁白,整個系列維持一致的角色音色。

  • 文字轉口播影片流程

    在 Hermes Agent 工作流程中把 ElevenLabs 語音與虛擬人模型串接,從文字直接得到有旁白的影片。

為甚麼透過 RunAPI + Hermes Agent 使用 ElevenLabs

  • 6 個變體,一個 API 金鑰

    透過一個 RunAPI 連線選擇可用的模型變體,無需更改現有整合。

  • 清晰的用量價格

    發送請求前查看目錄中的目前價格,無需訂閱或最低消費。

  • 自動化工作流程

    透過一致的工作流程提交工作、查詢狀態並收集非同步結果,無需編寫手動輪詢程式碼。

Hermes Agent + ElevenLabs 常見問題

可以在 Hermes Agent 中使用 ElevenLabs 嗎?

可以。在 Hermes Agent 中將 RunAPI 設定為供應商,再調用本頁列出的任一 ElevenLabs 端點,包括語音合成、轉錄、對話、音效和人聲分離。

可以在 Hermes Agent 中用 ElevenLabs 轉錄音訊嗎?

可以。用音訊 URL 調用語音轉文字端點。它支援區分說話者並標註音訊事件,結果以非同步方式回傳。

Hermes Agent 能把 ElevenLabs 與影片生成串接嗎?

可以。Hermes Agent 可以先用 ElevenLabs 生成語音,再在同一次執行中把音訊 URL 交給虛擬人或語音轉影片模型。

應該使用哪個模型 ID?

從版本表選擇公開模型 ID。每個 ID 可用的端點均列於相應資料列。

本指南會設定聊天模型嗎?

不會。此 Model Line 使用頁面所示的端點流程,不會被描述為 Agent 聊天模型。

透過 Hermes Agent 開始使用 ElevenLabs

正在與團隊一起建立?

我們可以協助企業設定、整合和技術問題。

聯絡我們