It looks like you may prefer a different language. Switch anytime.

LLM API-gateways

Een LiteLLM-alternatief dat je niet hoeft te hosten

LiteLLM is open-source software you deploy on your own server — Docker, config files, provider keys, uptime your problem. RunAPI is a hosted gateway: same OpenAI-compatible API, same 100+ LLMs, no server to run. Change the base URL and your existing code keeps working. RunAPI also adds image, video, and music generation, built-in billing, MCP server, CLI, and SDKs at 50% off official rates.

Bijgewerkt op June 23, 2026 RunAPI Editorial
In het kort

LiteLLM vs RunAPI — zelf-gehost of beheerd?

LiteLLM is an open-source Python proxy you run on your own server. It supports 100+ LLMs through a unified OpenAI-compatible interface — you pay provider costs and your own hosting, but you control everything. RunAPI is the managed version of that idea: the same OpenAI-compatible API, no server to run, 50% off official model rates, and coverage extended to image, video, and music generation under one API key. If the setup and ops work of self-hosting isn't something you want to own, RunAPI gets you the same multi-model access without it.

Hostingmodel

LiteLLM draait op jouw infrastructuur (Docker, VM's, K8s). RunAPI is een gehoste service — stuur verzoeken naar api.runapi.ai, geen inzet vereist.

Modaliteitsdekking

LiteLLM routeert tekst-LLM's. RunAPI routeert LLM's en voegt beeld-, video-, muziek- en audiogeneratie toe met één sleutel.

Prijzen

LiteLLM is gratis software; je betaalt aanbieders-API-kosten plus hosting. RunAPI rekent 50% van de officiële modeltarieven zonder infrastructuurkosten.

Ontwikkelaarstoolkit

RunAPI wordt geleverd met MCP-server, CLI en SDK's voor Python, Node.js, PHP, Java, Ruby en Go naast volledige OpenAI-SDK-compatibiliteit.

Vergelijking naast elkaar

Functieverlijking LiteLLM vs RunAPI

The table compares the two gateways across the dimensions that matter most: whether you want to run your own server or just use an API, how much you pay, and what models you can actually reach.

Functie LiteLLM RunAPI
Hostingmodel Zelf-gehost (jouw infrastructuur) Volledig beheerde service
LLM-routering Ja — 100+ modellen Ja — 100+ modellen
Beeldgeneratie Nee Ja
Videogeneratie Nee Ja
Muziek en audio Nee Ja
OpenAI-compatibele API Ja Ja
Ingebouwde facturering Handmatig (eigen sleutels) Ja — uniforme accountfacturering
MCP-server Nee Ja
Officiële SDK's Python (LiteLLM SDK) Python, Node.js, PHP, Java, Ruby, Go + OpenAI-compatibel
CLI litellm CLI (serverbeheer) Ja — generatie en sleutelbeheer
Kosten Aanbiederskosten + hosting 50% korting op officiële aanbiederstarieven
Zelf-gehost vs beheerd

Wat verandert er als je de zelf-gehoste proxy weggooit?

Running LiteLLM means cloning the repo, editing a .env file with a master key and salt key, running docker-compose, then configuring each provider through the admin UI — and that's just the first time. After that you own the uptime, the upgrades, and the key rotation. RunAPI removes that layer entirely. Send requests to api.runapi.ai and pay per call. No server, no config files, no provider accounts to juggle.

Geen server om in te richten

LiteLLM requires a running Python process or Docker container on a VPS you pay for and monitor. RunAPI is an HTTPS endpoint — point your client at it and start calling.

Geen configuratiebestanden om te onderhouden

De config.yaml van LiteLLM koppelt modelnamen aan aanbiederssleutels en verwerkt fallbacks. RunAPI verwerkt routering intern; jij levert één API-sleutel.

Geen sleutelbeheer per aanbieder

Met LiteLLM sla je Anthropic-, OpenAI-, Gemini- en andere sleutels op je server op. RunAPI bewaart één sleutel aan jouw kant — het beheert aanbiedersreferenties aan zijn kant.

Uniforme facturering op één plek

LiteLLM gives you raw provider invoices across multiple accounts. RunAPI consolidates all usage into a single dashboard with per-call cost records — one bill instead of a bunch of paid API accounts.

Modeldekking

Welke modellen routeert RunAPI?

LiteLLM's strength is breadth — it routes 100+ LLMs from every major provider. RunAPI matches that LLM coverage across the frontier models most teams actually use (Claude, GPT, Gemini, DeepSeek) and adds generative media models not available through LiteLLM.

Taalmodellen

Claude, GPT, Gemini, DeepSeek, and other LLMs for chat, coding, reasoning, and tool use — the same text coverage LiteLLM users expect from the frontier models.

Beeldmodellen

Tekst-naar-beeld- en beeldbewerkingsmodellen, waaronder Flux en Seedream, aanroepbaar via dezelfde API-sleutel als je LLM-aanroepen.

Videomodellen

Tekst-naar-video- en beeld-naar-videomodellen voor het genereren van korte clips, gefactureerd op basis van betaling per taak.

Muziek en audio

Muziekgeneratie en audiomodellen voor soundtracks en stemwerk — modaliteiten volledig buiten het bereik van LiteLLM.

Prijsvergelijking

Hoe verhoudt de prijsstelling van RunAPI zich tot zelf-hosting van LiteLLM?

LiteLLM is free software but you pay full provider rates on top of what your VPS costs (~$6–12/month) plus time spent maintaining the setup. RunAPI charges 50% of official provider rates with no hosting overhead. For teams that aren't optimizing LiteLLM's advanced routing features, the total cost often favors the managed option.

Model Officiële invoer /M Officiële uitvoer /M RunAPI invoer /M RunAPI uitvoer /M
Claude Sonnet 4.6 $6,00 $30,00 $3,00 $15,00
Claude Opus 4.7 $10,00 $50,00 $5,00 $25,00
GPT-5.4 $2,50 $15,00 $1,25 $7,50
Gemini 2.5 Pro $1,25 $10,00 $0,63 $5,00

RunAPI past een korting van 50% toe op alle LLM-aanbieders. Mediamodellen worden per taak gefactureerd. Prijzen geverifieerd juni 2026.

Aan de slag

Hoe schakel je over van LiteLLM naar RunAPI

1

Maak een RunAPI-account aan

Meld je aan op runapi.ai. Er is geen creditcard vereist om te starten.

2

Haal je API-sleutel op

Ga naar Dashboard -> API-sleutels, maak een sleutel aan en kopieer hem naar je omgeving.

3

Werk de basis-URL bij

Wijzig in je OpenAI-compatibele client of LiteLLM-configuratie de basis-URL naar https://api.runapi.ai/v1 en vervang je LiteLLM-sleutel door de RunAPI-sleutel. Er zijn geen andere codewijzigingen nodig.

4

Decommissioneer je LiteLLM-server

Zodra verzoeken via RunAPI worden gerouteerd, kun je je LiteLLM-proces stoppen en de server verwijderen. Aanbiederssleutels die daar zijn opgeslagen, zijn niet langer nodig.

Veelgestelde vragen

Veelgestelde vragen over LiteLLM-alternatief

Isn't this still paying for cloud APIs? How is this different from just using the services directly?

Good question — the difference is one API key, one bill, and one base URL. With LiteLLM you still need an Anthropic key, an OpenAI key, a Gemini key, and so on, plus the LiteLLM server to hold them all together. RunAPI gives you a single key that reaches every model. You drop it in where you used to put your OpenAI key, change the base URL to api.runapi.ai/v1, and you're done. No accounts to manage per provider, no invoices per service.

You started with self-hosted, then immediately added a bunch of paid cloud API subscriptions — isn't that the opposite of self-hosted?

LiteLLM gives you self-hosting with full control — your server, your provider keys, your config files. RunAPI is the other side of that choice: no Docker, no VPS to pay for, no config.yaml to maintain, no uptime to monitor. You still get the same OpenAI-compatible API and access to 100+ LLMs. The tradeoff is that you don't run the infrastructure; RunAPI does. If control and privacy are the priority, LiteLLM makes sense. If you want multi-model access without the setup, RunAPI gets you there faster.

Is RunAPI a drop-in replacement for LiteLLM?

RunAPI is a drop-in replacement for the OpenAI-compatible API surface — the same /v1/chat/completions endpoint. Change the base URL and API key in your existing client and requests start routing through RunAPI. LiteLLM-specific features like virtual keys and custom routing rules have no equivalent, so check those before switching.

Hoeveel kost RunAPI vergeleken met LiteLLM?

LiteLLM is free software but you pay full provider rates plus the cost of running a server (a VPS typically runs $6–12/month) plus time spent on upgrades, config changes, and debugging outages. RunAPI charges 50% of official model rates with no hosting overhead. For Claude Sonnet 4.6, that is $3 per million input tokens and $15 per million output tokens versus $6 and $30 at official rates — and nothing extra for the server.

Werkt RunAPI met Claude Code, Cursor en Windsurf?

Ja. Stel de Anthropic- of OpenAI-basis-URL in elke tool in op het RunAPI-endpoint en gebruik je RunAPI-sleutel. Alle drie de tools werken zonder aanvullende aanpassingen. Dezelfde sleutel ontgrendelt ook beeld-, video- en audiomodellen voor projecten die multimodale generatie nodig hebben.

Wat is de MCP-server en hoe verschilt die van LiteLLM?

RunAPI wordt geleverd met een Model Context Protocol-server waarmee MCP-bewuste hosts zoals Claude Code modellen kunnen ontdekken, prijzen kunnen controleren en taken rechtstreeks vanuit de assistentinterface kunnen aanmaken. LiteLLM heeft geen MCP-server. De RunAPI MCP-server is aanvullend — je kunt hem naast de REST API gebruiken in plaats van ervoor.

Kan ik de OpenAI Python SDK met RunAPI gebruiken?

Yes. Point the base_url parameter at https://api.runapi.ai/v1 and pass your RunAPI key as api_key. The rest of your OpenAI SDK code stays unchanged. RunAPI also publishes its own typed SDKs for Python, Node.js, PHP, Java, Ruby, and Go if you prefer a client built for RunAPI's full API surface including media generation.

Wat gebeurt er met LiteLLM-functies zoals virtuele sleutels en uitgavenregistratie?

LiteLLM's virtual key system lets you issue internal keys with per-key spend caps and model restrictions — useful if you're running LiteLLM as a shared hub for a family or team. RunAPI doesn't have virtual keys, but provides per-account usage reporting and per-call cost records in the dashboard. If the main reason you use LiteLLM's virtual keys is to track or limit spend, RunAPI's account billing covers that. If you need to issue keys to other people with different model access, evaluate that gap before migrating.

GERELATEERD

Meer ontdekken

Het gehoste alternatief voor het uitvoeren van je eigen proxy.

RunAPI geeft je 130+ LLM- en mediamodellen met 50% korting op de officiële tarieven — geen server om uit te voeren, geen configuratie om te onderhouden, één sleutel om te beheren.