Text · Z.ai

GLM API

Z.ai GLM API access via RunAPI — MIT-licensed MoE models with up to 200K context, leading open-weight coding benchmarks.

جاهز للتشغيل · 7 variants · ابتداءً من $0.010

جرّبه في مساحة التجربة مرجع API →

runapi.ai

# Base URL
https://runapi.ai

# Endpoints
POST /v1/chat/completions

curl https://runapi.ai/v1/chat/completions \
  -H "Authorization: Bearer $RUNAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "glm-5.1",
  "messages": [
    {
      "role": "user",
      "content": "Read this multi-file repository, find the failing integration test, and propose a patch with an explanation of the root cause."
    }
  ]
}'

from openai import OpenAI

client = OpenAI(
    base_url="https://runapi.ai/v1",
    api_key="your-runapi-key"
)

response = client.chat.completions.create(
    model="glm-5.1",
    messages=[{"role": "user", "content": "Read this multi-file repository, find the failing integration test, and propose a patch with an explanation of the root cause."}]
)

import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://runapi.ai/v1",
  apiKey: "your-runapi-key"
});

const response = await client.chat.completions.create({
  model: "glm-5.1",
  messages: [{ role: "user", content: "Read this multi-file repository, find the failing integration test, and propose a patch with an explanation of the root cause." }]
});

https://runapi.ai /v1/chat/completions

OVERVIEW

GLM is Z.ai's family of MIT-licensed Mixture-of-Experts language models. GLM-4.5 (355B total / 32B active, 128K context) introduced the open-weight MoE line with a flagship and a lighter Air tier. GLM-4.6 and 4.7 extend to 200K context with stronger code generation — 4.7 reaches 73.8% on SWE-bench. The GLM-5 series (744B / 40B active, 200K context) pushes further to 77.8% SWE-bench Verified, and GLM-5.1 holds the top open-weight score on SWE-bench Pro at 58.4%. All are available through RunAPI with one key and per-token billing.

عدة إصدارات لمستويات مختلفة من السرعة / الجودة
يتضمن model skill التوثيق والـschemas وملاحظات الإعداد
يعمل مع workflows موجهة لتطوير التطبيقات
لا يتم احتساب رسوم على عمليات التوليد الفاشلة

VARIANTS

قارن جميع إصدارات API

Variant	Billing	From
glm-4.5	1K tokens	$0.020	عرض →
glm-4.5-air	1K tokens	$0.010	عرض →
glm-4.6	1K tokens	$0.020	عرض →
glm-4.7	1K tokens	$0.020	عرض →
glm-5	1K tokens	$0.020	عرض →
glm-5-turbo	1K tokens	$0.020	عرض →
glm-5.1	1K tokens	$0.030	عرض →

API

نقاط نهاية API لـ GLM

استخدم SDK OpenAI أو Anthropic مع مفتاح RunAPI. لا حاجة لـ SDK إضافي.

Endpoint	Protocol
/v1/chat/completions	OpenAI compatible

كيف يعمل

من model skill إلى أول نتيجة في أربع خطوات

اختر النموذج

اختر النموذج والإصدار المناسبين لنوع المخرجات ومستوى الجودة وزمن الاستجابة المستهدف.

الإعداد

اضبط مفتاح RunAPI وثبّت model skill في مساحة عمل الكود.

التنفيذ

استخدم تعليمات المهارة لإضافة ميزة النموذج داخل تطبيقك.

الاستلام

استعلم عبر task ID أو استخدم البث عند دعمه أو عالج webhook callback.

CONTEXT

ما هي GLM API؟

GLM models from Z.ai are MIT-licensed MoE LLMs spanning 128K–200K context. GLM-5.1 leads open-weight models on SWE-bench Pro. Through RunAPI they share a single API key with pay-as-you-go token billing, callable from the OpenAI Chat Completions, OpenAI Responses, and Anthropic Messages surfaces.

Provider

Z.ai

Modality

Text

لماذا RunAPI