Back to models/MiniMax-M2.5
稀宇科技
MiniMaxCallable

MiniMax-M2.5

Great for customer support, content production, knowledge-base Q&A, and tool-chain agents.

Model code
MiniMax-M2.5
TextLong contextTool use

Context

256K tokens

Availability

1/1 available

Reference latency

2.50s

Capabilities

Inferred from the model family and tags; actual calls are authoritative.

Function calling
Supported
Structured output
Not supported
Vision
Not supported
Image generation
Not supported
Web search
Not supported
Code execution
Not supported
Streaming
Supported
Caching
Supported
Batch inference
Not supported

Pricing

Final cost is determined at settlement.

Input$0.34 / M
Output$1.37 / M
Cache hitHit $0.07 / M
Pricing statusUnified pricing

Pricing comes from the unified pricing config; balance and budget are checked before each call.

Limits & context

Rate-limit fields sync from the console config; no guarantees until synced.

Max context
200K tokens
Max output
200K tokens
RPM
No hard cap · fair use
TPM
Tiered by account & key

API examples

Use Turiloop's unified API entry; read the key from an environment variable.

import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.TURILOOP_API_KEY,
  baseURL: "https://api.turing.yun/v1"
});

const completion = await client.chat.completions.create({
  model: "MiniMax-M2.5",
  messages: [{ role: "user", content: "Hello, Turiloop" }]
});