Public Versions and Endpoints
| Model ID | Endpoints | Price | Catalog |
|---|---|---|---|
audio-isolation
|
/api/v1/elevenlabs/isolate_audio
|
$0.12 / minute | Model detail |
sound-effect-v2
|
/api/v1/elevenlabs/text_to_sound
|
$0.15 / minute | Model detail |
speech-to-text
|
/api/v1/elevenlabs/speech_to_text
|
$0.04 / minute | Model detail |
text-to-dialogue-v3
|
/api/v1/elevenlabs/text_to_dialogue
|
$0.14 / 1K chars | Model detail |
text-to-speech-multilingual-v2
|
/api/v1/elevenlabs/text_to_speech
|
$0.12 / 1K chars | Model detail |
text-to-speech-turbo-v2.5
|
/api/v1/elevenlabs/text_to_speech
|
$0.06 / 1K chars | Model detail |
Verify
Poll until the task reaches a terminal status
Select <model-id> to generate verification commands.
Configuration
Guide endpoint: <endpoint>
Select <model-id> to generate a request with the endpoint's public input contract.
Get Started in 3 Steps
-
Choose a model ID
Select a public catalog model ID and review its endpoint and current starting price.
-
Configure RunAPI
Set RUNAPI_API_KEY before making the endpoint request.
-
Verify the result
For asynchronous endpoints, poll the same endpoint until the Task reaches a terminal status.
What to Build with OpenClaw + ElevenLabs
-
Audiobook and podcast narration
Convert long-form text into spoken audio with consistent character voices across hours of content.
-
Video dubbing into multiple languages
Dub video content into other languages with the same voice profile, producing localized versions that keep the speaker's vocal character.
-
Sound effects for video and games
Generate custom Foley, ambient audio, and sound cues from text descriptions instead of searching stock audio libraries.
Why Use ElevenLabs Through RunAPI + OpenClaw
-
6 variants, one API key
Use one RunAPI connection to choose among the live model variants without changing your integration.
-
Clear usage pricing
See current catalog pricing before you send a request, with no subscription or minimum spend required.
-
Automatic task workflows
Submit, poll, and collect asynchronous results through a consistent task workflow without writing manual polling code.
OpenClaw + ElevenLabs Questions
Which voice settings sound most natural?
Higher stability makes a voice more consistent but less expressive, while higher similarity keeps it closer to the original profile. Use more stability for narration and less for conversational content.
How should I handle long-form content such as audiobooks?
Split long text into chunks and send one request per chunk, processing chapters in parallel instead of one after another.
Can I create multi-speaker audio in OpenClaw?
Yes. Call the dialogue endpoint with a list of lines, each paired with a voice, and ElevenLabs returns one conversation with several speakers.
Which model ID should I use?
Choose a public model ID from the version table. Each ID exposes the endpoints shown for that version.
Does this guide configure a chat model?
No. This Model Line uses the endpoint workflow shown here and is not presented as an agent chat model.