Public Versions and Endpoints
| Model ID | Endpoints | Price | Catalog |
|---|---|---|---|
qwen3.5-122b-a10b
|
/v1/chat/completions
|
$0.40 / 1M tokens | Model detail |
qwen3.5-27b
|
/v1/chat/completions
|
$0.30 / 1M tokens | Model detail |
qwen3.5-35b-a3b
|
/v1/chat/completions
|
$0.25 / 1M tokens | Model detail |
qwen3.5-397b-a17b
|
/v1/chat/completions
|
$0.60 / 1M tokens | Model detail |
qwen3.6-27b
|
/v1/chat/completions
|
$0.60 / 1M tokens | Model detail |
qwen3.6-35b-a3b
|
/v1/chat/completions
|
$0.38 / 1M tokens | Model detail |
Verify
Poll until the task reaches a terminal status
Select <model-id> to generate verification commands.
Configuration
Guide endpoint: <endpoint>
Select <model-id> to generate a request with the endpoint's public input contract.
Get Started in 3 Steps
-
Choose a model ID
Select a public catalog model ID and review its endpoint and current starting price.
-
Configure RunAPI
Add the provider configuration for this agent.
-
Verify the result
Run the agent status and model-selection commands shown below.
What to Build with Hermes Agent + Qwen
-
Agent coding and reasoning
Let your agent plan, write, and review code or work through multi-step problems with the model's reasoning.
-
Long document analysis and extraction
Summarize reports, contracts, or codebases and pull structured fields out of long documents.
-
Tool-calling workflows
Connect the model to functions and APIs so your agent can look up data and take actions between replies.
Why Use Qwen Through RunAPI + Hermes Agent
-
6 variants, one API key
Use one RunAPI connection to choose among the live model variants without changing your integration.
-
Clear usage pricing
See current catalog pricing before you send a request, with no subscription or minimum spend required.
-
Direct responses
Synchronous calls return the result in the same response, so your agent can use it immediately without task polling.
Hermes Agent + Qwen Questions
Which Qwen model should I start with?
Start with qwen3.6-27b or qwen3.6-35b-a3b, the newest Qwen models on RunAPI, for chat, coding, and image input. Use qwen3.5-397b-a17b when you want the largest Qwen3.5 model, and qwen3-coder-next for coding agents that do not need a thinking step.
Which Qwen models accept images?
All Qwen3.5 and Qwen3.6 models and both Qwen3-VL-235B-A22B versions accept images. Send them as image_url parts in a Chat Completions request. qwen3-coder-next is text-only.
How does thinking work on Qwen models?
The Qwen3.5 and Qwen3.6 models are hybrid and have thinking and non-thinking modes. qwen3-vl-235b-a22b-thinking always thinks, while qwen3-vl-235b-a22b-instruct and qwen3-coder-next answer without a thinking step.
Which API do I use to call Qwen on RunAPI?
Use the OpenAI-compatible Chat Completions endpoint at https://runapi.ai/v1/chat/completions with your RunAPI API key and a Qwen model ID. Any OpenAI SDK works after you change the base URL.
Are the Qwen language models open-weight?
Yes. Alibaba releases these Qwen models with open weights under the Apache 2.0 license. RunAPI gives you API access, so you do not have to host them on your own GPUs.
How is the Qwen line different from Qwen Image?
The Qwen line holds the language models, which read text and images and return text. Qwen Image, Qwen 2, and Qwen 3 are separate RunAPI lines that generate and edit images.
Which model ID should I use?
Choose a public model ID from the version table. Each ID exposes the endpoints shown for that version.
Does this guide configure a chat model?
Yes. This Model Line supports a client-facing LLM protocol.