Model Catalog

Explore 249 AI Models

Browse, compare, and integrate the best AI models for video, image, music, audio, and text generation — all through one unified API.

All models available · Real-time pricing
# Available Endpoints
POST /v1/chat/completions # LLM
POST /v1/images/generations # Image
POST /api/v1/kling/text_to_video # Video
POST /v1/audio/speech # TTS
POST /v1/audio/transcriptions # STT
POST /api/v1/suno/text_to_music # Music

Leading AI models, one API

  • Flux
  • Midjourney
  • Whisper
  • ElevenLabs
  • HappyHorse
  • GPT Image 2
  • Veo
  • Minimax
  • Luma
  • PixVerse
  • Recraft
  • Fish Audio
  • DeepSeek
  • Qwen
  • Claude
  • GPT
  • Gemini
  • Seedance
  • Kling
  • Suno
Provider

56 models

Text
Claude
Anthropic

Claude API access for Anthropic's LLM across complex reasoning, code, analysis, and extended-context tasks.

Text
DeepSeek
DeepSeek

DeepSeek API access via RunAPI — flash for fast, low-cost work; pro for complex agentic tasks.

Audio & Music
ElevenLabs
ElevenLabs

ElevenLabs API access for voice synthesis, text-to-speech, sound effects, speech-to-text, and audio isolation.

Text
Embedding
OpenAI

OpenAI text embeddings for semantic search, retrieval, clustering, and ranking workflows.

Audio & Music
Fish Audio
Fish Audio

Fish Audio API access for expressive multilingual and production-grade text-to-speech with managed MP3 or WAV output.

Image
Flux
Black Forest Labs

Flux image generation with Dev, Pro, and 2 Klein, plus single-image editing with Dev and Pro.

Image
Flux 2
Black Forest Labs

Flux 2 API access for text-to-image and remix-image with strong prompt adherence from Black Forest Labs.

Image
Flux Kontext
Black Forest Labs

Flux Kontext API access for in-context image editing, local edits, style transfer, and character consistency.

Text
Gemini
Google

Gemini API access for Google's multimodal LLM across chat, code generation, reasoning, and long-context tasks.

Video
Gemini Omni
Google

Gemini Omni API access for voice, character, and multimodal video resources in agent media workflows.

Audio & Music
Gemini TTS
Google

Gemini TTS API access for multi-speaker dialogue with configurable voices, accents, delivery styles, and pacing.

Text
GLM
Z.ai

Z.ai GLM API access via RunAPI — MIT-licensed MoE models with up to 200K context, leading open-weight coding benchmarks.

Text
GPT
OpenAI

GPT API access for OpenAI's flagship LLM across chat, code generation, and multi-step reasoning tasks.

Image
GPT Image
OpenAI

GPT Image API access for text-to-image and image editing powered by OpenAI image generation models.

Image
GPT Image 2
OpenAI

GPT Image 2 API access for OpenAI image generation with near-perfect multilingual text rendering.

Image
GPT Image 2.5
OpenAI

Use the GPT Image 2.5 API via RunAPI with a model skill, unified auth, and pay-as-you-go pricing.

Image
GPT-4o Image
OpenAI

GPT-4o Image API access for native image generation and editing inside the conversation.

Text
Grok
xAI

xAI Grok 4.5 and 4.6 API access through RunAPI Responses and Chat Completions for reasoning, tools, and structured output.

Image
Grok Imagine
xAI

Grok Imagine API access for image and video generation from text, including text-to-image, image-to-image, text-to-video, and image-to-video.

Video
Grok Imagine Video
xAI

Use the Grok Imagine Video API via RunAPI with a model skill, unified auth, and pay-as-you-go pricing.

Video
Hailuo
MiniMax

Hailuo API access for text and image-to-video at native 1080p with accurate physics simulation and motion.

Video
HappyHorse
Alibaba

HappyHorse API access for text, image, and edit-video generation with 720p/1080p output and character-guided clips.

Image
Ideogram V3
Ideogram

Ideogram V3 API access for text-to-image with strong in-image text accuracy for posters, logos, and typography.

Image
Imagen 4
Google

Imagen 4 API access for photorealistic text-to-image, precise typography, and broad styles.

Video
InfiniteTalk
MeiGen-AI

InfiniteTalk API access for audio-driven talking-head animation, lip-sync, and portrait animation.

JE Utility
Jev
TypeSafe

TypeSafe Jev API for typed decisions in software — classify, route, score, and branch with calibrated confidence.

Text
Kimi
Moonshot AI

Moonshot AI Kimi API access via RunAPI — kimi-k3 is the current flagship with always-on reasoning; kimi-k2.5 and kimi-k2.6 remain available.

Video
Kling
Kuaishou

Kling API access for text and image-to-video up to 4K with multimodal audio and AI avatars.

Video
Luma
Luma

Luma API access for video modification and transformation powered by Dream Machine.

Image
Midjourney Image
Midjourney

Use the Midjourney Image API via RunAPI with a model skill, unified auth, and pay-as-you-go pricing.

Video
Midjourney Video
Midjourney

Use the Midjourney Video API via RunAPI with a model skill, unified auth, and pay-as-you-go pricing.

Text
MiMo
Xiaomi

MiMo API via RunAPI — base mimo-v2.5 supports text and remote HTTP(S) images in synchronous OpenAI Chat Completions; mimo-v2.5-pro remains text-only.

Text
MiniMax
MiniMax

MiniMax text API access via RunAPI — sparse MoE models from 200K to 1M context, up to 80.5% SWE-bench Verified.

Video
MiniMax H3
MiniMax

MiniMax H3 video API for 2K text-to-video, image-to-video, and multi-reference generation from images, video, and audio.

Image
Nano Banana
Google

Nano Banana API access for fast text-to-image with accurate in-image text and multi-character consistency.

Utility
OmniHuman
Bytedance

OmniHuman API access for audio-driven talking-head video, human identification, and subject-mask detection.

Utility
OpenAI Moderation
OpenAI

Use the OpenAI Moderation API via RunAPI with a model skill, unified auth, and pay-as-you-go pricing.

Audio & Music
OpenAI Transcription
OpenAI

OpenAI-compatible audio transcription API access for speech-to-text with Whisper-1 and GPT Transcribe.

Audio & Music
OpenAI TTS
OpenAI

OpenAI TTS API access for low-latency and high-quality text-to-speech with RunAPI-managed MP3 output.

Video
PixVerse
PixVerse

PixVerse V6 video API via RunAPI — generate video from text, images, reference sets, frame transitions, and extend existing results.

Audio & Music
Producer
Producer

Producer API access for FUZZ music generation from exact lyrics or instrumental production briefs.

Image
Qwen 2
Alibaba

Qwen 2 API access for text-to-image and image editing from Alibaba's visual model family.

Image
Qwen 3
Alibaba

Qwen 3 image generation and editing in standard and Pro tiers.

Image
Qwen Image
Alibaba

Qwen Image API access for high-quality image generation, prompt-guided remixing, and precise text-driven editing.

Utility
Recraft
Recraft

Recraft API access for AI image upscaling and background removal in design and production workflows.

Video
Runway
Runway

Runway API access for video generation and editing — create and transform footage with text prompts.

Video
Runway Aleph
Runway

Runway Aleph API access for prompt-guided video editing with frame-level continuity.

Video
Seedance
Bytedance

Seedance API access for text and image-to-video with native audio-video synthesis and multi-shot clips.

Image
Seedream
Bytedance

Seedream API access for text-to-image and image editing with strong typography rendering and up to 4K resolution.

Audio & Music
Suno
Suno

Suno API access for AI music generation — create songs with vocals, instruments, and lyrics from a text prompt.

Utility
Topaz
Topaz

Topaz Labs API access for AI image and video upscaling that enhances resolution and detail.

Video
Veo 3.1
Google

Veo 3.1 API access for high-fidelity video generation up to 4K with synthesized dialogue, sound effects, and ambience.

Video
Volcengine Lip Sync
Bytedance

Volcengine Lip Sync API access for audio-driven video-to-video lip sync from source video and audio.

Image
Wan Image
Alibaba

Use the Wan Image API via RunAPI with a model skill, unified auth, and pay-as-you-go pricing.

Video
Wan Video
Alibaba

Use the Wan Video API via RunAPI with a model skill, unified auth, and pay-as-you-go pricing.

Image
Z Image
Alibaba

Z Image API access for ultra-fast text-to-image, photorealistic results, and short inference paths.

56 models

Try a Model Now

Pick any model and generate in seconds.

Estimated: $0.61 / second

Sign in to generate with RunAPI.

Sign in to generate

Sign in to start generating with your available language models.

WHY RUNAPI

One API, Every AI Model

Unified Interface

One SDK for 240+ models across all modalities. Switch providers without changing code.

Real-time Pricing

Transparent pay-as-you-go pricing. No subscriptions, no hidden fees, credits never expire.

Enterprise Ready

Team API keys, usage analytics, and priority support.

Get Started in Seconds

Python SDK
pip install runapi-suno
CLI
curl -fsSL https://runapi.ai/cli/install.sh | sh
MCP Server
npx -y @runapi.ai/mcp

What teams build with 200+ models

Video Marketing

Generate ad variants per audience segment with Kling, Veo, and Runway

Product Photography

Create studio-quality product shots with Flux and Midjourney

Podcast Production

Clone voices and generate intros with Fish Audio and ElevenLabs

AI Agents

Give your agent access to any model through MCP or REST

Music for Content

Score videos and podcasts with Suno and Udio

Multilingual TTS

Localize apps into 40+ languages with natural voices

Catalog questions

How do I browse available models?

Visit /models or use the CLI: runapi models list. Filter by modality, provider, or search by name.

Can I use multiple models in one project?

Yes. One API key works across all 240+ models. Switch between video, image, music, audio, and LLM models freely.

How are new models added?

We add new models continuously. Major releases are typically available within days of launch. Follow our changelog for updates.

Can I request a specific model?

Yes. Contact us with the model name and provider. We evaluate all requests and prioritize based on demand.

What happens when a model is retired?

We notify users 30 days before retirement. Your existing code continues working — just update the model parameter to a newer version.

How are models priced?

Text models bill per million input and output tokens; media models bill per second, per image, or per minute. Every rate is listed on the model page and in the pricing table.

Start building with 249 models.