OpenAI · Audio & Music

OpenClaw x OpenAI Transcription

OpenAI Transcription provides speech-to-text models for meeting notes, captions, searchable media, and multilingual transcripts. Both variants use the OpenAI-compatible audio transcription request shape.

Public versions and endpoints

Model IDVersionEndpointsStarting priceCatalog
gpt-transcribeMultilingual speech-to-text with keyword and language hints
/v1/audio/transcriptions
$0.020Model detail
whisper-1Flexible audio transcription with subtitle and timestamp output
/v1/audio/transcriptions
$0.020Model detail

Configuration

Guide endpoint: <endpoint>

Select <model-id> to generate a request with the endpoint's public input contract.

Verify

Select <model-id> to generate verification commands.

How it works

  1. 1. Choose a model ID

    Select a public catalog model ID and review its endpoint and current starting price.

  2. 2. Configure RunAPI

    Set RUNAPI_API_KEY before making the endpoint request.

  3. 3. Verify the result

    For asynchronous endpoints, poll the same endpoint until the Task reaches a terminal status.

FAQ

Which model ID should I use?

Choose a public model ID from the version table. Each ID exposes the endpoints shown for that version.

Does this guide configure a chat model?

No. This Model Line uses the endpoint workflow shown here and is not presented as an agent chat model.