Kein Server bereitzustellen
LiteLLM requires a running Python process or Docker container on a VPS you pay for and monitor. RunAPI is an HTTPS endpoint — point your client at it and start calling.
LiteLLM is open-source software you deploy on your own server — Docker, config files, provider keys, uptime your problem. RunAPI is a hosted gateway: same OpenAI-compatible API, same 100+ LLMs, no server to run. Change the base URL and your existing code keeps working. RunAPI also adds image, video, and music generation, built-in billing, MCP server, CLI, and SDKs at 50% off official rates.
LiteLLM is an open-source Python proxy you run on your own server. It supports 100+ LLMs through a unified OpenAI-compatible interface — you pay provider costs and your own hosting, but you control everything. RunAPI is the managed version of that idea: the same OpenAI-compatible API, no server to run, 50% off official model rates, and coverage extended to image, video, and music generation under one API key. If the setup and ops work of self-hosting isn't something you want to own, RunAPI gets you the same multi-model access without it.
LiteLLM läuft auf Ihrer Infrastruktur (Docker, VMs, K8s). RunAPI ist ein gehosteter Dienst — senden Sie Anfragen an runapi.ai, keine Bereitstellung erforderlich.
LiteLLM routet Text-LLMs. RunAPI routet LLMs und fügt Bild-, Video-, Musik- und Audiogenerierung unter einem Schlüssel hinzu.
LiteLLM ist kostenlose Software; Sie zahlen Provider-API-Kosten plus Hosting. RunAPI berechnet 50 % der offiziellen Modellraten ohne Infrastruktur-Overhead.
RunAPI liefert einen MCP-Server, CLI und SDKs für Python, Node.js, PHP, Java, Ruby und Go neben vollständiger OpenAI-SDK-Kompatibilität.
The table compares the two gateways across the dimensions that matter most: whether you want to run your own server or just use an API, how much you pay, and what models you can actually reach.
| Funktion | LiteLLM | RunAPI |
|---|---|---|
| Hosting-Modell | Selbst gehostet (eigene Infrastruktur) | Vollständig verwalteter Dienst |
| LLM-Routing | Ja — 100+ Modelle | Ja — 100+ Modelle |
| Bildgenerierung | Nein | Ja |
| Videogenerierung | Nein | Ja |
| Musik und Audio | Nein | Ja |
| OpenAI-kompatible API | Ja | Ja |
| Integrierte Abrechnung | Manuell (eigene Schlüssel) | Ja — einheitliche Kontoabrechnung |
| MCP-Server | Nein | Ja |
| Offizielle SDKs | Python (LiteLLM SDK) | Python, Node.js, PHP, Java, Ruby, Go + OpenAI-kompatibel |
| CLI | litellm CLI (Server-Management) | Ja — Generierung und Schlüsselverwaltung |
| Kosten | Provider-Kosten + Hosting | 50 % Rabatt auf offizielle Provider-Tarife |
Running LiteLLM means cloning the repo, editing a .env file with a master key and salt key, running docker-compose, then configuring each provider through the admin UI — and that's just the first time. After that you own the uptime, the upgrades, and the key rotation. RunAPI removes that layer entirely. Send requests to runapi.ai and pay per call. No server, no config files, no provider accounts to juggle.
LiteLLM requires a running Python process or Docker container on a VPS you pay for and monitor. RunAPI is an HTTPS endpoint — point your client at it and start calling.
LiteLLMs config.yaml ordnet Modellnamen Provider-Schlüsseln zu und verwaltet Fallbacks. RunAPI verwaltet das Routing intern; Sie stellen einen API-Schlüssel bereit.
Mit LiteLLM speichern Sie Anthropic-, OpenAI-, Gemini- und andere Schlüssel auf Ihrem Server. RunAPI hält einen Schlüssel auf Ihrer Seite — es verwaltet Provider-Zugangsdaten auf seiner Seite.
LiteLLM gives you raw provider invoices across multiple accounts. RunAPI consolidates all usage into a single dashboard with per-call cost records — one bill instead of a bunch of paid API accounts.
LiteLLM's strength is breadth — it routes 100+ LLMs from every major provider. RunAPI matches that LLM coverage across the frontier models most teams actually use (Claude, GPT, Gemini, DeepSeek) and adds generative media models not available through LiteLLM.
Claude, GPT, Gemini, DeepSeek, and other LLMs for chat, coding, reasoning, and tool use — the same text coverage LiteLLM users expect from the frontier models.
Text-zu-Bild- und Bildbearbeitungsmodelle einschließlich Flux und Seedream, über denselben API-Schlüssel wie Ihre LLM-Aufrufe abrufbar.
Text-zu-Video- und Bild-zu-Video-Modelle zur Generierung kurzer Clips, abgerechnet nach Nutzung pro Aufgabe.
Musikgenerierung und Audiomodelle für Soundtracks und Spracharbeit — Modalitäten, die außerhalb des Anwendungsbereichs von LiteLLM liegen.
LiteLLM is free software but you pay full provider rates on top of what your VPS costs (~$6–12/month) plus time spent maintaining the setup. RunAPI charges 50% of official provider rates with no hosting overhead. For teams that aren't optimizing LiteLLM's advanced routing features, the total cost often favors the managed option.
| Modell | Offizielle Eingabe /M | Offizielle Ausgabe /M | RunAPI Eingabe /M | RunAPI Ausgabe /M |
|---|---|---|---|---|
| Claude Sonnet 4.6 | $6,00 | $30,00 | $3,00 | $15,00 |
| Claude Opus 4.7 | $10,00 | $50,00 | $5,00 | $25,00 |
| GPT-5.4 | $2,50 | $15,00 | $1,25 | $7,50 |
| Gemini 2.5 Pro | $1,25 | $10,00 | $0,63 | $5,00 |
RunAPI gewährt 50 % Rabatt bei allen LLM-Providern. Medienmodelle werden pro Aufgabe abgerechnet. Preise verifiziert Juni 2026.
Melden Sie sich bei runapi.ai an. Keine Kreditkarte erforderlich, um loszulegen.
Gehen Sie zu Dashboard -> API Keys, erstellen Sie einen Schlüssel und kopieren Sie ihn in Ihre Umgebung.
Ändern Sie in Ihrem OpenAI-kompatiblen Client oder Ihrer LiteLLM-Konfiguration die Base URL auf https://runapi.ai/v1 und ersetzen Sie Ihren LiteLLM-Schlüssel durch den RunAPI-Schlüssel. Keine weiteren Codeänderungen sind erforderlich.
Sobald Anfragen über RunAPI geroutet werden, können Sie Ihren LiteLLM-Prozess stoppen und den Server entfernen. Die dort gespeicherten Provider-Schlüssel werden nicht mehr benötigt.
Good question — the difference is one API key, one bill, and one base URL. With LiteLLM you still need an Anthropic key, an OpenAI key, a Gemini key, and so on, plus the LiteLLM server to hold them all together. RunAPI gives you a single key that reaches every model. You drop it in where you used to put your OpenAI key, change the base URL to runapi.ai/v1, and you're done. No accounts to manage per provider, no invoices per service.
LiteLLM gives you self-hosting with full control — your server, your provider keys, your config files. RunAPI is the other side of that choice: no Docker, no VPS to pay for, no config.yaml to maintain, no uptime to monitor. You still get the same OpenAI-compatible API and access to 100+ LLMs. The tradeoff is that you don't run the infrastructure; RunAPI does. If control and privacy are the priority, LiteLLM makes sense. If you want multi-model access without the setup, RunAPI gets you there faster.
RunAPI is a drop-in replacement for the OpenAI-compatible API surface — the same /v1/chat/completions endpoint. Change the base URL and API key in your existing client and requests start routing through RunAPI. LiteLLM-specific features like virtual keys and custom routing rules have no equivalent, so check those before switching.
LiteLLM is free software but you pay full provider rates plus the cost of running a server (a VPS typically runs $6–12/month) plus time spent on upgrades, config changes, and debugging outages. RunAPI charges 50% of official model rates with no hosting overhead. For Claude Sonnet 4.6, that is $3 per million input tokens and $15 per million output tokens versus $6 and $30 at official rates — and nothing extra for the server.
Ja. Setzen Sie die Anthropic- oder OpenAI-Base-URL in den Einstellungen jedes Tools auf den RunAPI-Endpunkt und verwenden Sie Ihren RunAPI-Schlüssel. Alle drei Tools funktionieren ohne zusätzliche Änderungen. Derselbe Schlüssel schaltet auch Bild-, Video- und Audiomodelle für Projekte frei, die multimodale Generierung benötigen.
RunAPI liefert einen Model Context Protocol-Server, mit dem MCP-fähige Hosts wie Claude Code Modelle entdecken, Preise prüfen und Aufgaben direkt über die Assistenten-Oberfläche erstellen können. LiteLLM hat keinen MCP-Server. Der RunAPI-MCP-Server ist additiv — Sie können ihn zusammen mit der REST-API verwenden, nicht statt ihr.
Yes. Point the base_url parameter at https://runapi.ai/v1 and pass your RunAPI key as api_key. The rest of your OpenAI SDK code stays unchanged. RunAPI also publishes its own typed SDKs for Python, Node.js, PHP, Java, Ruby, and Go if you prefer a client built for RunAPI's full API surface including media generation.
LiteLLM's virtual key system lets you issue internal keys with per-key spend caps and model restrictions — useful if you're running LiteLLM as a shared hub for a family or team. RunAPI doesn't have virtual keys, but provides per-account usage reporting and per-call cost records in the dashboard. If the main reason you use LiteLLM's virtual keys is to track or limit spend, RunAPI's account billing covers that. If you need to issue keys to other people with different model access, evaluate that gap before migrating.
RunAPI vs. Requesty: Preise, Funktionen und Modellauswahl im Vergleich.
RunAPI vs. OpenRouter: Preise, Modelle und Medien-Support im Vergleich.
Claude Opus, Sonnet und Haiku zu 50% der Anthropic-Listenpreise.
Alle 210+ Modelle mit Preisen, Parametern und Code-Beispielen durchsuchen.
RunAPI gibt Ihnen 210+ LLM- und Medienmodelle zu 50 % der offiziellen Tarife — kein Server zu betreiben, keine Konfiguration zu pflegen, ein Schlüssel zu verwalten.