Video · Google

Gemini Omni API

에이전트 미디어 워크플로우에서 음성, 캐릭터, 멀티모달 비디오 리소스를 위한 Gemini Omni API 액세스.

운영 중 · 4 variants · 최저 $0.0000
runapi.ai
curl -X POST https://runapi.ai/api/v1/gemini_omni/create_audio \
  -H "Authorization: Bearer $RUNAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "audio_id": "achernar",
  "name": "Acher Narrator",
  "voice_description": "침착한 나레이터가 말하는 동안 재사용 가능한 캐릭터가 빗속을 걷는 1080p 네온 도시 추적 영상을 만들어 주세요.",
  "example_dialogue": "Hello, I am achernar"
}'
import { GeminiOmniClient } from "@runapi.ai/gemini-omni";

const client = new GeminiOmniClient();
const result = await client.createAudio.run({
    audio_id: "achernar",
    name: "Acher Narrator",
    voice_description: "침착한 나레이터가 말하는 동안 재사용 가능한 캐릭터가 빗속을 걷는 1080p 네온 도시 추적 영상을 만들어 주세요.",
    example_dialogue: "Hello, I am achernar",
});
<?php

require __DIR__ . "/vendor/autoload.php";

use RunApi\GeminiOmni\GeminiOmniClient;

$client = new GeminiOmniClient();
$result = $client->createAudio->run([
        'audio_id' => 'achernar',
        'name' => 'Acher Narrator',
        'voice_description' => '침착한 나레이터가 말하는 동안 재사용 가능한 캐릭터가 빗속을 걷는 1080p 네온 도시 추적 영상을 만들어 주세요.',
        'example_dialogue' => 'Hello, I am achernar',
]);
require "runapi/gemini_omni"

client = RunApi::GeminiOmni::Client.new
result = client.create_audio.run(
    audio_id: "achernar",
    name: "Acher Narrator",
    voice_description: "침착한 나레이터가 말하는 동안 재사용 가능한 캐릭터가 빗속을 걷는 1080p 네온 도시 추적 영상을 만들어 주세요.",
    example_dialogue: "Hello, I am achernar"
)
npx skills add runapi-ai/gemini-omni -g
# Claude Code
claude mcp add runapi -s user -- npx -y @runapi.ai/mcp

# Codex
codex plugin install runapi-mcp@agents

# Cursor / Windsurf / VS Code
npx @runapi.ai/mcp init cursor
@runapi.ai/gemini-omni v1
개요

Gemini Omni는 재사용 가능한 음성 리소스, 재사용 가능한 캐릭터 리소스, 그리고 프롬프트, 참조 이미지, 오디오 ID, 캐릭터 ID, 소스 비디오 클립을 결합할 수 있는 멀티모달 비디오를 생성합니다.

  • 품질과 지연 시간 목표에 맞춘 모델 변형
  • 통합 API key
  • Model skill에 docs, schema, 설정 메모 포함
  • 실패한 생성은 과금되지 않습니다
변형

변형

Variant Billing From
gemini-omni-audio call $0.0000 보기 →
gemini-omni-character call $0.0000 보기 →
gemini-omni-flash-preview call $0.600 보기 →
gemini-omni-text-to-video call $3.60 보기 →
MODELS

앱 개발을 위해 Gemini Omni skill 설치

모델 docs, schema, 가격 메모, 설정 단계를 코딩 워크스페이스로 불러옵니다.

# Install the model skill for app development workflows
npx skills add runapi-ai/gemini-omni -g
Installs docs, schemas, pricing context, and setup notes into your developer workspace.
Or use this setup request in your coding tool:
Install the Gemini Omni skill for this app:

1. Add runapi-ai/gemini-omni with the skills installer.
2. Load SKILL.md in this workspace.
3. Use its docs, schemas, pricing notes, and setup steps when adding model features.
4. Confirm the install path when done.
작동 방식

이 model skill로 구현하는 방법

01

모델 선택

출력 유형, 품질 기준, 지연 시간 목표에 맞는 모델과 변형을 고릅니다.

02

한 번 인증

모든 지원 모델에 RunAPI key를 사용합니다.

03

skill 설치

기능을 구현하기 전에 코딩 워크스페이스에 model skill을 추가합니다.

04

결과 받기

task ID로 조회하거나 생성 완료 시 callback을 처리합니다.

컨텍스트

Gemini Omni의 위치

Gemini Omni는 RunAPI의 Google 카탈로그에 속하며, 오디오, 캐릭터, 비디오 변형 전반에 걸쳐 동일한 SDK 패키지, CLI 네임스페이스, 과금 방식을 공유합니다.

Provider
Google
Modality
Video
RUNAPI를 선택하는 이유

RunAPI로 Gemini Omni을 쓰는 이유

하나의 API key

모델과 제공사를 넘나들며 같은 인증 정보를 사용합니다.

Skill-ready

model skill에 schema, 설정 메모, 가격 컨텍스트, 모델 ID가 포함됩니다.

예측 가능한 과금

호출 전에 사용량 기반 가격을 확인할 수 있습니다.

FAQ

자주 묻는 질문

이 모델은 어떻게 호출하나요?

model skill을 설치하고 RunAPI key와 함께 설정 메모를 따르세요.

실패한 생성도 비용이 드나요?

실패한 생성은 과금되지 않습니다

애플리케이션에서 호출할 수 있나요?

네. 코딩 워크스페이스에 model skill을 설치하고 모델 기능을 추가할 때 사용하세요.

지금 시작

Gemini Omni로 개발을 시작하세요.