It looks like you may prefer a different language. Switch anytime.

Video Google

Gemini Omni API

透過 RunAPI 使用 Gemini Omni API,包含 model skill、統一驗證與按用量計費。

runapi.ai
curl -X POST https://runapi.ai/api/v1/gemini_omni/create_audio \
  -H "Authorization: Bearer $RUNAPI_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "audio_id": "achernar",
  "name": "Acher Narrator",
  "voice_description": "Create a 1080p neon city tracking shot with a reusable character walking through rain while a calm narrator speaks.",
  "example_dialogue": "Hello, I am achernar"
}'
import { GeminiOmniClient } from "@runapi.ai/gemini-omni";

const client = new GeminiOmniClient();
const result = await client.createAudio.run({
    audio_id: "achernar",
    name: "Acher Narrator",
    voice_description: "Create a 1080p neon city tracking shot with a reusable character walking through rain while a calm narrator speaks.",
    example_dialogue: "Hello, I am achernar",
});
<?php

require __DIR__ . "/vendor/autoload.php";

use RunApi\GeminiOmni\GeminiOmniClient;

$client = new GeminiOmniClient();
$result = $client->createAudio->run([
        'audio_id' => 'achernar',
        'name' => 'Acher Narrator',
        'voice_description' => 'Create a 1080p neon city tracking shot with a reusable character walking through rain while a calm narrator speaks.',
        'example_dialogue' => 'Hello, I am achernar',
]);
require "runapi/gemini_omni"

client = RunApi::GeminiOmni::Client.new
result = client.create_audio.run(
    audio_id: "achernar",
    name: "Acher Narrator",
    voice_description: "Create a 1080p neon city tracking shot with a reusable character walking through rain while a calm narrator speaks.",
    example_dialogue: "Hello, I am achernar"
)
npx skills add runapi-ai/gemini-omni -g
# Claude Code
claude mcp add runapi -s user -- npx -y @runapi.ai/mcp

# Codex
codex plugin install runapi-mcp@agents

# Cursor / Windsurf / VS Code
npx @runapi.ai/mcp init cursor
@runapi.ai/gemini-omni v1
OVERVIEW

關於 Gemini Omni

Gemini Omni creates reusable voice resources, reusable character resources, and multimodal videos that can combine prompts, reference images, audio IDs, character IDs, and a source video clip.

供應商
Google
模態
Video

比較所有 API 變體

變體 計費方式 定價
gemini-omni-audio call $0.0000 查看 →
gemini-omni-character call $0.0000 查看 →
gemini-omni-flash-1-1 call $2.52 查看 →
gemini-omni-flash-preview call $0.460 查看 →
gemini-omni-text-to-video call $2.52 查看 →
PRICING

Gemini Omni 定價

Endpoint 解析度 時長 價格
create_audio Free / track

顯示 gemini-omni-audio 的定價。其他變體保有各自的 endpoint 定價。

Agent 整合

透過 Agent 執行 Gemini Omni

運作方式

Gemini Omni 快速入門

  1. 選擇模型

    挑選符合輸出類型、品質門檻與延遲目標的模型與變體。

  2. 設定

    設定 RunAPI key,並在 coding workspace 安裝 model skill。

  3. 開發

    使用 skill 指引,在你的 app 內加入模型功能。

  4. 接收

    透過 task ID 查詢、在支援時串流,或處理 webhook callback。

CONTEXT

什麼是 Gemini Omni API?

Gemini Omni belongs to the Google catalog on RunAPI and shares the same SDK package, CLI namespace, and billing surfaces across audio, character, and video variants.

供應商
Google
查看全部 →
模態
Video
瀏覽模型 →

為什麼透過 RunAPI 使用 Gemini Omni API

一組驗證,全部供應商通用

一把 RunAPI 金鑰就能開通整個目錄。無需分開註冊帳戶,也不用為每個整合各自輪替金鑰。

統一價格與計費

以美元按次計費,每月結帳。失敗的生成不收費。

內含 schema 的 skill

型別化 schema 與 setup 備註打包在 model skill 內,讓實作從正確契約開始。

Gemini Omni 常見問題

What can Gemini Omni create?

It can create reusable audio resources, reusable character resources, and multimodal videos that combine prompts with image, audio, character, or source-video references.

When should I create a character resource?

Create a character resource when multiple videos need the same visual identity; pass the returned character ID into video requests that should reuse that subject.

How do audio resources work?

Audio resources capture a preset voice choice and can be referenced by video requests that need narration or dialogue with consistent voice selection.

What affects Gemini Omni video pricing?

Video price depends on the request mode, duration, resolution, and whether source media changes the effective billed clip length.

Can one video request mix several reference types?

Yes. Gemini Omni video can combine prompt text with reference images, audio IDs, character IDs, and a source video, within the documented reference limits.

我應該先從哪個版本開始?

先選擇符合你品質標準中最便宜的版本。大多數團隊會先用快速版本,之後再升級到專業版用於正式上線。

相似模型

如果你喜歡 Gemini Omni API,可以試試這些

開始用 Gemini Omni API 開發。