OpenAI · Audio & Music
Hermes Agent x OpenAI Transcription
OpenAI Transcription provides speech-to-text models for meeting notes, captions, searchable media, and multilingual transcripts. Both variants use the OpenAI-compatible audio transcription request shape.
Public versions and endpoints
| Model ID | Version | Endpoints | Starting price | Catalog |
|---|---|---|---|---|
gpt-transcribe | Multilingual speech-to-text with keyword and language hints | /v1/audio/transcriptions | $0.020 | Model detail |
whisper-1 | Flexible audio transcription with subtitle and timestamp output | /v1/audio/transcriptions | $0.020 | Model detail |
Configuration
Guide endpoint: <endpoint>
Select <model-id> to generate a request with the endpoint's public input contract.Verify
Select <model-id> to generate verification commands.How it works
- 1. Choose a model ID
Select a public catalog model ID and review its endpoint and current starting price.
- 2. Configure RunAPI
Set RUNAPI_API_KEY before making the endpoint request.
- 3. Verify the result
For asynchronous endpoints, poll the same endpoint until the Task reaches a terminal status.
FAQ
Which model ID should I use?
Choose a public model ID from the version table. Each ID exposes the endpoints shown for that version.
Does this guide configure a chat model?
No. This Model Line uses the endpoint workflow shown here and is not presented as an agent chat model.