gpt-transcribe API
OpenAI / OpenAI Transcription
透過 RunAPI 使用 OpenAI Transcription 系列的 gpt-transcribe。按次收費,無需訂閱,失敗的生成不扣費。
可正常運作
·
audio_music
·
可商用
curl -X POST https://runapi.ai/v1/audio/transcriptions \
-H "Authorization: Bearer $RUNAPI_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gpt-transcribe",
"audio_url": "https://cdn.runapi.ai/public/samples/voice.mp3"
}'
import { OpenaiTranscriptionClient } from "@runapi.ai/openai-transcription";
const client = new OpenaiTranscriptionClient();
const result = await client.speechToText.run({
model: "gpt-transcribe",
audio_url: "https://cdn.runapi.ai/public/samples/voice.mp3",
});
require "runapi/openai_transcription"
client = RunApi::OpenaiTranscription::Client.new
result = client.speech_to_text.run(
model: "gpt-transcribe",
audio_url: "https://cdn.runapi.ai/public/samples/voice.mp3"
)
npx skills add runapi-ai/openai-transcription -g
# Claude Code
claude mcp add runapi -s user -- npx -y @runapi.ai/mcp
# Codex
codex plugin install runapi-mcp@agents
# Cursor / Windsurf / VS Code
npx @runapi.ai/mcp init cursor
切換 variant
OVERVIEW
gpt-transcribe 針對 OpenAI Transcription 系列中品質與成本的最佳平衡點。
- 按次收費,美元計價
- 生成失敗不收費
- 如模型支援,提供串流輸出
- Model skill setup
PRICING
收費
失敗的生成不會收費
Speech to text
$0.02
/ minute
規格說明
技術詳情
| 模型 ID | gpt-transcribe |
| 供應商 | OpenAI |
| 模式 | audio_music |
| 任務類型 | synchronous |
| 計費單位 | minute |
| API endpoint | /v1/audio/transcriptions |
| 商業授權 | 是 — 已透過 API 包含 |
| 目錄狀態 | 可正常運作 |
SKILLS
快速開始 — gpt-transcribe
結構一致 · variant 已固定於 model 中
# Install the model skill for app development workflows
npx skills add runapi-ai/openai-transcription -g
Or use this setup request in your coding tool:
Install the OpenAI Transcription skill for this app: 1. Add runapi-ai/openai-transcription with the skills installer. 2. Load SKILL.md in this workspace. 3. Use its docs, schemas, pricing notes, and setup steps when adding model features. 4. Confirm the install path when done.
運作方式
四步使用 gpt-transcribe
01
安裝
為此模型系列安裝 model skill。
02
設定
將 model 欄位設為本頁顯示的完整 model ID。
03
呼叫
連同 prompt、inputs 及 callback 設定,送出一個 typed request。
04
接收
從 RunAPI 讀取 task response、webhook callback 或 cached output URL。
DIFFERENCES
gpt-transcribe 有咩不同
VS WHISPER-1
支援關鍵字同語言提示嘅多語言語音轉文字
靈活音訊轉錄,支援字幕同時間戳輸出
使用場景
最適合
Podcast 及影片配樂
生成同集數氛圍匹配的免版稅背景音樂,無需付授權費。
遊戲音效
為程序生成關卡製作可自適應的環境音景同音效。
廣告旁白及音效
無需錄音室,都可以為客戶廣告生成自訂旁白同音效。
FAQ
關於 gpt-transcribe 的常見問題
模型 ID 會否在不同版本之間保持穩定?
RunAPI 會保持 model ID 穩定,並處理相容的版本更新,而不需要改動你的 request 格式。
這個 variant 的 rate limit 係多少?
每個 key 的 rate limit 會按使用層級而定。請查看定價頁了解最新限制。
之後可以切換 variant 嗎?
可以——variant 只係一個旗標。只要更改 model 參數就可以切換。
支援 streaming 嗎?
在可用 streaming 的情況下,RunAPI 會端到端串流。
我應該在邊度回報品質問題?
可以在公開 GitHub repo 開 issue,或者電郵支援團隊。
OpenAI Transcription 的其他 variant
其他模型的替代方案
立即開始