ElevenLabs · 音訊與音樂

Hermes Agent x ElevenLabs

ElevenLabs 是語音 AI 公司,其模型涵蓋 TTS、對話、音效、轉錄與音訊分離。透過 RunAPI,所有 ElevenLabs 端點共用一組金鑰,並按次計費。

6 variants 起價 $0.04 / 分鐘 可商用

Prerequisite: npx runapi mcp install

Prompt

Prompt 模型

任務合適時使用 audio-isolation:從混合音源中擷取人聲。

You have access to RunAPI task tools for ElevenLabs.

Available ElevenLabs models:
- audio-isolation: 從混合音源中擷取人聲
  endpoints: /api/v1/elevenlabs/isolate_audio
  request fields: source_audio_url
- sound-effect-v2: 文字生成音效,適用於遊戲、影片與 Podcast
  endpoints: /api/v1/elevenlabs/text_to_sound
  request fields: text, loop, duration_seconds, prompt_influence, output_format
- speech-to-text: 支援 29 種以上語言的轉錄,並具備說話者分段
  endpoints: /api/v1/elevenlabs/speech_to_text
  request fields: source_audio_url, language_code, diarize
- text-to-dialogue-v3: 多位說話者的對話生成,自然輪流發言
  endpoints: /api/v1/elevenlabs/text_to_dialogue
  request fields: dialogue, stability, language_code
- text-to-speech-multilingual-v2: 29 種語言;最逼真的情感表現;非常適合有聲書
  endpoints: /api/v1/elevenlabs/text_to_speech
  request fields: model, text, voice, language_code
- text-to-speech-turbo-v2.5: 32 種語言;非英語快 3 倍;40K 字元上限
  endpoints: /api/v1/elevenlabs/text_to_speech
  request fields: model, text, voice, language_code

Use the model ID and endpoint that match the user's request.
Example prompt: 將這段文字轉換成自然語音,使用溫暖的英國男聲,語速適中。
/api/v1/elevenlabs/isolate_audio: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/isolate_audio,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/isolate_audio/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/elevenlabs/text_to_sound: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/text_to_sound,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/text_to_sound/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/elevenlabs/speech_to_text: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/speech_to_text,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/speech_to_text/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/elevenlabs/text_to_dialogue: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/text_to_dialogue,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/text_to_dialogue/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。
/api/v1/elevenlabs/text_to_speech: Submit the task, poll for its status, then verify the completed output. POST /api/v1/elevenlabs/text_to_speech,並儲存回傳的 Task ID。
GET /api/v1/elevenlabs/text_to_speech/<task-id>,直到狀態為 completed 或 failed。
確認 completed 回應包含 API 參考中記載的端點專屬輸出。

公開版本與端點

模型 ID 端點 起始價格 模型目錄
audio-isolation
/api/v1/elevenlabs/isolate_audio
$0.12 / 分鐘 模型詳情
sound-effect-v2
/api/v1/elevenlabs/text_to_sound
$0.15 / 分鐘 模型詳情
speech-to-text
/api/v1/elevenlabs/speech_to_text
$0.04 / 分鐘 模型詳情
text-to-dialogue-v3
/api/v1/elevenlabs/text_to_dialogue
$0.14 / 1K 字元 模型詳情
text-to-speech-multilingual-v2
/api/v1/elevenlabs/text_to_speech
$0.12 / 1K 字元 模型詳情
text-to-speech-turbo-v2.5
/api/v1/elevenlabs/text_to_speech
$0.06 / 1K 字元 模型詳情

驗證

輪詢,直到 Task 進入終態

選擇 <model-id> 後產生驗證命令。

設定

指南端點: <endpoint>

選擇 <model-id> 後,依端點的公開輸入契約產生請求。
使用流程

三步開始使用

  1. 選擇模型 ID

    選擇公開模型 ID,並查看其端點與目前起始價格。

  2. 設定 RunAPI

    呼叫端點前設定 RUNAPI_API_KEY。

  3. 驗證結果

    對非同步端點輪詢同一路徑,直到 Task 進入終態。

使用 Hermes Agent + ElevenLabs 可以打造什麼

  • 對話式語音 Agent

    建立說話自然的語音 Agent,以低延遲語音支援客服機器人、助理或電話介面。

  • YouTube 內容旁白

    為 YouTube 影片製作旁白,整個系列維持一致的角色音色。

  • 文字轉口播影片流程

    在 Hermes Agent 工作流程中把 ElevenLabs 語音與虛擬人模型串接,從文字直接得到有旁白的影片。

為什麼透過 RunAPI + Hermes Agent 使用 ElevenLabs

  • 6 個變體,一個 API 金鑰

    透過一個 RunAPI 連線選擇可用的模型變體,無需變更現有整合。

  • 清楚的用量價格

    送出請求前查看目錄中的目前價格,無需訂閱或最低消費。

  • 自動化工作流程

    透過一致的工作流程提交工作、查詢狀態並收集非同步結果,無需撰寫手動輪詢程式碼。

Hermes Agent + ElevenLabs 常見問題

可以在 Hermes Agent 中使用 ElevenLabs 嗎?

可以。在 Hermes Agent 中將 RunAPI 設定為提供者,再呼叫本頁列出的任一 ElevenLabs 端點,包括語音合成、轉錄、對話、音效和人聲分離。

可以在 Hermes Agent 中用 ElevenLabs 轉錄音訊嗎?

可以。用音訊 URL 呼叫語音轉文字端點。它支援區分說話者並標註音訊事件,結果以非同步方式回傳。

Hermes Agent 能把 ElevenLabs 與影片產生串接嗎?

可以。Hermes Agent 可以先用 ElevenLabs 產生語音,再在同一次執行中把音訊 URL 交給虛擬人或語音轉影片模型。

應該使用哪個模型 ID?

從版本表選擇公開模型 ID。每個 ID 可用的端點都列在對應資料列。

本指南會設定聊天模型嗎?

不會。此 Model Line 使用頁面所示的端點流程,不會被描述為 Agent 聊天模型。

開始在 Hermes Agent 中使用 ElevenLabs

正在與團隊一起打造?

我們可以協助企業接入、整合與技術問題。

聯絡我們