Public Versions and Endpoints
| Model ID | Endpoints | Price | Catalog |
|---|---|---|---|
MiniMax-M2
|
/v1/chat/completions
|
$0.19 / 1M tokens | Model detail |
MiniMax-M2.1
|
/v1/chat/completions
|
$0.19 / 1M tokens | Model detail |
MiniMax-M2.5
|
/v1/chat/completions
|
$0.19 / 1M tokens | Model detail |
MiniMax-M2.5-highspeed
|
/v1/chat/completions
|
$0.37 / 1M tokens | Model detail |
MiniMax-M2.7
|
/v1/chat/completions
|
$0.19 / 1M tokens | Model detail |
MiniMax-M2.7-highspeed
|
/v1/chat/completions
|
$0.37 / 1M tokens | Model detail |
MiniMax-M3
|
/v1/chat/completions
|
$0.18 / 1M tokens | Model detail |
Verify
Poll until the task reaches a terminal status
Select <model-id> to generate verification commands.
Configuration
Guide endpoint: <endpoint>
Select <model-id> to generate a request with the endpoint's public input contract.
Get Started in 3 Steps
-
Choose a model ID
Select a public catalog model ID and review its endpoint and current starting price.
-
Configure RunAPI
Add the provider configuration for this agent.
-
Verify the result
Run the agent status and model-selection commands shown below.
What to Build with Hermes Agent + MiniMax
-
Agent coding and reasoning
Let your agent plan, write, and review code or work through multi-step problems with the model's reasoning.
-
Long document analysis and extraction
Summarize reports, contracts, or codebases and pull structured fields out of long documents.
-
Tool-calling workflows
Connect the model to functions and APIs so your agent can look up data and take actions between replies.
Why Use MiniMax Through RunAPI + Hermes Agent
-
7 variants, one API key
Use one RunAPI connection to choose among the live model variants without changing your integration.
-
Clear usage pricing
See current catalog pricing before you send a request, with no subscription or minimum spend required.
-
Direct responses
Synchronous calls return the result in the same response, so your agent can use it immediately without task polling.
Hermes Agent + MiniMax Questions
Are these the same as MiniMax Hailuo video models?
No. These are MiniMax's text language models for coding and chat; Hailuo is MiniMax's separate video generation line.
Which MiniMax text model should I pick?
MiniMax-M3 is the strongest — 1M context, 80.5% SWE-bench Verified, and the first open-weight model to combine frontier coding with million-token context. M2.7 is the best 200K-context option. Highspeed variants (M2.5 and M2.7) run the same weights at ~100 tokens/sec for lower latency at higher token cost.
Which SDKs can call MiniMax text through RunAPI?
Use the OpenAI SDK (Chat Completions or Responses) or the Anthropic Messages SDK against RunAPI with the MiniMax model id; the proxy adapts the protocol.
How is MiniMax text billed?
Per token at RunAPI's published input and output rates for each model, pay-as-you-go. Highspeed variants are billed at their own published rates for the same output quality at higher throughput.
Which model ID should I use?
Choose a public model ID from the version table. Each ID exposes the endpoints shown for that version.
Does this guide configure a chat model?
Yes. This Model Line supports a client-facing LLM protocol.