프로비저닝할 서버 없음
LiteLLM requires a running Python process or Docker container on a VPS you pay for and monitor. RunAPI is an HTTPS endpoint — point your client at it and start calling.
LiteLLM is open-source software you deploy on your own server — Docker, config files, provider keys, uptime your problem. RunAPI is a hosted gateway: same OpenAI-compatible API, same 100+ LLMs, no server to run. Change the base URL and your existing code keeps working. RunAPI also adds image, video, and music generation, built-in billing, MCP server, CLI, and SDKs at 50% off official rates.
LiteLLM is an open-source Python proxy you run on your own server. It supports 100+ LLMs through a unified OpenAI-compatible interface — you pay provider costs and your own hosting, but you control everything. RunAPI is the managed version of that idea: the same OpenAI-compatible API, no server to run, 50% off official model rates, and coverage extended to image, video, and music generation under one API key. If the setup and ops work of self-hosting isn't something you want to own, RunAPI gets you the same multi-model access without it.
LiteLLM은 직접 인프라(Docker, VM, K8s)에서 실행됩니다. RunAPI는 호스팅 서비스로, runapi.ai에 요청을 보내면 배포가 필요하지 않습니다.
LiteLLM은 텍스트 LLM을 라우팅합니다. RunAPI는 LLM을 라우팅하며 하나의 키로 이미지, 비디오, 음악, 오디오 생성을 추가합니다.
LiteLLM은 무료 소프트웨어이며, 공급자 API 비용과 호스팅 비용을 지불합니다. RunAPI는 인프라 오버헤드 없이 공식 모델 요금의 50%를 청구합니다.
RunAPI는 완전한 OpenAI SDK 호환성과 함께 Python, Node.js, PHP, Java, Ruby, Go용 MCP 서버, CLI, SDK를 제공합니다.
The table compares the two gateways across the dimensions that matter most: whether you want to run your own server or just use an API, how much you pay, and what models you can actually reach.
| 기능 | LiteLLM | RunAPI |
|---|---|---|
| 호스팅 방식 | 자체 호스팅 (직접 인프라) | 완전 관리형 서비스 |
| LLM 라우팅 | 가능 — 100개 이상 모델 | 가능 — 100개 이상 모델 |
| 이미지 생성 | 불가 | 가능 |
| 비디오 생성 | 불가 | 가능 |
| 음악 및 오디오 | 불가 | 가능 |
| OpenAI 호환 API | 가능 | 가능 |
| 내장 청구 | 수동 (자체 키 사용) | 가능 — 통합 계정 청구 |
| MCP 서버 | 불가 | 가능 |
| 공식 SDK | Python (LiteLLM SDK) | Python, Node.js, PHP, Java, Ruby, Go + OpenAI 호환 |
| CLI | litellm CLI (서버 관리) | 가능 — 생성 및 키 관리 |
| 비용 | 공급자 비용 + 호스팅 | 공식 공급자 요금의 50% 할인 |
Running LiteLLM means cloning the repo, editing a .env file with a master key and salt key, running docker-compose, then configuring each provider through the admin UI — and that's just the first time. After that you own the uptime, the upgrades, and the key rotation. RunAPI removes that layer entirely. Send requests to runapi.ai and pay per call. No server, no config files, no provider accounts to juggle.
LiteLLM requires a running Python process or Docker container on a VPS you pay for and monitor. RunAPI is an HTTPS endpoint — point your client at it and start calling.
LiteLLM의 config.yaml은 모델 이름을 공급자 키에 매핑하고 폴백을 처리합니다. RunAPI는 내부적으로 라우팅을 처리하며, API 키 하나만 제공하면 됩니다.
LiteLLM을 사용하면 Anthropic, OpenAI, Gemini 및 기타 키를 서버에 저장합니다. RunAPI는 사용자 측에서 하나의 키만 보유하며, 공급자 자격 증명은 RunAPI 측에서 관리합니다.
LiteLLM gives you raw provider invoices across multiple accounts. RunAPI consolidates all usage into a single dashboard with per-call cost records — one bill instead of a bunch of paid API accounts.
LiteLLM's strength is breadth — it routes 100+ LLMs from every major provider. RunAPI matches that LLM coverage across the frontier models most teams actually use (Claude, GPT, Gemini, DeepSeek) and adds generative media models not available through LiteLLM.
Claude, GPT, Gemini, DeepSeek, and other LLMs for chat, coding, reasoning, and tool use — the same text coverage LiteLLM users expect from the frontier models.
Flux와 Seedream을 포함한 텍스트-이미지 및 이미지 편집 모델로, LLM 호출과 동일한 API 키로 호출할 수 있습니다.
짧은 클립을 생성하기 위한 텍스트-비디오 및 이미지-비디오 모델로, 작업당 종량제로 청구됩니다.
사운드트랙과 음성 작업을 위한 음악 생성 및 오디오 모델 — LiteLLM의 범위를 완전히 벗어난 모달리티입니다.
LiteLLM is free software but you pay full provider rates on top of what your VPS costs (~$6–12/month) plus time spent maintaining the setup. RunAPI charges 50% of official provider rates with no hosting overhead. For teams that aren't optimizing LiteLLM's advanced routing features, the total cost often favors the managed option.
| 모델 | 공식 입력 /M | 공식 출력 /M | RunAPI 입력 /M | RunAPI 출력 /M |
|---|---|---|---|---|
| Claude Sonnet 4.6 | $6.00 | $30.00 | $3.00 | $15.00 |
| Claude Opus 4.7 | $10.00 | $50.00 | $5.00 | $25.00 |
| GPT-5.4 | $2.50 | $15.00 | $1.25 | $7.50 |
| Gemini 2.5 Pro | $1.25 | $10.00 | $0.63 | $5.00 |
RunAPI는 모든 LLM 공급자에 50% 할인을 적용합니다. 미디어 모델은 작업당 청구됩니다. 2026년 6월 기준 가격.
runapi.ai에서 가입합니다. 시작하는 데 신용카드가 필요하지 않습니다.
Dashboard -> API Keys로 이동하여 키를 생성하고 환경에 복사합니다.
OpenAI 호환 클라이언트나 LiteLLM 구성에서 Base URL을 https://runapi.ai/v1으로 변경하고 LiteLLM 키를 RunAPI 키로 교체합니다. 다른 코드 변경은 필요하지 않습니다.
요청이 RunAPI를 통해 라우팅되면 LiteLLM 프로세스를 중지하고 서버를 제거할 수 있습니다. 저장된 공급자 키는 더 이상 필요하지 않습니다.
Good question — the difference is one API key, one bill, and one base URL. With LiteLLM you still need an Anthropic key, an OpenAI key, a Gemini key, and so on, plus the LiteLLM server to hold them all together. RunAPI gives you a single key that reaches every model. You drop it in where you used to put your OpenAI key, change the base URL to runapi.ai/v1, and you're done. No accounts to manage per provider, no invoices per service.
LiteLLM gives you self-hosting with full control — your server, your provider keys, your config files. RunAPI is the other side of that choice: no Docker, no VPS to pay for, no config.yaml to maintain, no uptime to monitor. You still get the same OpenAI-compatible API and access to 100+ LLMs. The tradeoff is that you don't run the infrastructure; RunAPI does. If control and privacy are the priority, LiteLLM makes sense. If you want multi-model access without the setup, RunAPI gets you there faster.
RunAPI is a drop-in replacement for the OpenAI-compatible API surface — the same /v1/chat/completions endpoint. Change the base URL and API key in your existing client and requests start routing through RunAPI. LiteLLM-specific features like virtual keys and custom routing rules have no equivalent, so check those before switching.
LiteLLM is free software but you pay full provider rates plus the cost of running a server (a VPS typically runs $6–12/month) plus time spent on upgrades, config changes, and debugging outages. RunAPI charges 50% of official model rates with no hosting overhead. For Claude Sonnet 4.6, that is $3 per million input tokens and $15 per million output tokens versus $6 and $30 at official rates — and nothing extra for the server.
네. 각 도구의 설정에서 Anthropic 또는 OpenAI Base URL을 RunAPI 엔드포인트로 설정하고 RunAPI 키를 사용합니다. 세 도구 모두 추가 수정 없이 작동합니다. 동일한 키로 멀티모달 생성이 필요한 프로젝트를 위한 이미지, 비디오, 오디오 모델도 사용할 수 있습니다.
RunAPI는 Claude Code와 같은 MCP 지원 호스트가 모델을 검색하고, 요금을 확인하고, 어시스턴트 인터페이스에서 직접 작업을 생성할 수 있는 Model Context Protocol 서버를 제공합니다. LiteLLM에는 MCP 서버가 없습니다. RunAPI MCP 서버는 추가적이며, REST API 대신이 아닌 함께 사용할 수 있습니다.
Yes. Point the base_url parameter at https://runapi.ai/v1 and pass your RunAPI key as api_key. The rest of your OpenAI SDK code stays unchanged. RunAPI also publishes its own typed SDKs for Python, Node.js, PHP, Java, Ruby, and Go if you prefer a client built for RunAPI's full API surface including media generation.
LiteLLM's virtual key system lets you issue internal keys with per-key spend caps and model restrictions — useful if you're running LiteLLM as a shared hub for a family or team. RunAPI doesn't have virtual keys, but provides per-account usage reporting and per-call cost records in the dashboard. If the main reason you use LiteLLM's virtual keys is to track or limit spend, RunAPI's account billing covers that. If you need to issue keys to other people with different model access, evaluate that gap before migrating.
RunAPI는 공식 요금의 50%로 210+ LLM 및 미디어 모델을 제공합니다. 실행할 서버도, 유지할 구성도, 관리할 키도 하나뿐입니다.