SmartBrain API

Enterprise AI as Simple as Utilities

More Links
Model HubConsoleAPI KeysUsage QueryDocs
SmartBrain API. All rights reserved|Privacy Policy|Terms
SmartBrain API
SmartBrain API
  • Model Hub
  • Official Docs
  • Playground
M

minimax/MiniMax-M3

Online Chat
MiniMax

Publish time

2026/5/31

Model Series

MiniMax

Input type

Output type

Input Price

¥2.1 / 1M tokens

Output Price

¥8.4 / 1M tokens

Cache Write Price

¥2.625 / 1M tokens

Cache Read Price

¥0.42 / 1M tokens

Context Window

512,000

Max Output Length

128,000

MiniMax-M3 是 MiniMax 推出的多模态基础模型。它支持文本、图像和视频输入,并输出文本,拥有 1M 的上下文窗口,适用于长时间的智能体工作、编码和工具使用。该模型基于 MiniMax 稀疏注意力机制 (MSA) 构建,MSA 用键值块选择取代了完整的注意力机制,从而在长时间上下文中大幅降低每个词元的计算量——在 1M 个token的情况下,其计算成本约为上一代模型的 1/20,同时显著加快了预填充和解码速度,并在大多数任务中保持了质量。 该模型在交错数据上作为原生多模态模型进行训练,并通过交互式用户模拟器框架针对多轮次、类似生产环境的协作进行了调优,因此更适合持续的多步骤任务,而非单轮执行。

Providers for minimax/MiniMax-M3

Zhinao API routes requests to the best-fit provider and automatically fails over to the one with highest availability.

Sort by
M
MiniMax
国内

TTFT

2.47s

Throughput

41.12tps

Uptime

100.00%

Provider Model

minimax/minimax/MiniMax-M3

Supported Parameters

temperaturetop_ptop_k

Recent Uptime

10月11日 11 PM100.00%
10月9日 7 PM: 99.90%10月9日 7 PM: 99.60%10月9日 8 PM: 99.72%10月9日 8 PM: 100.00%10月9日 9 PM: 99.52%10月9日 9 PM: 100.00%10月9日 10 PM: 99.54%10月9日 10 PM: 100.00%10月9日 11 PM: 100.00%10月9日 11 PM: 99.75%10月10日 8 AM: 99.70%10月10日 9 AM: 100.00%10月10日 9 AM: 99.27%10月10日 10 AM: 99.85%10月10日 10 AM: 99.24%10月10日 11 AM: 99.74%10月10日 11 AM: 99.30%10月10日 12 PM: 99.64%10月10日 12 PM: 99.44%10月10日 1 PM: 99.66%10月10日 1 PM: 99.19%10月10日 2 PM: 99.74%10月10日 2 PM: 99.38%10月10日 3 PM: 99.51%10月10日 3 PM: 99.67%10月10日 4 PM: 99.85%10月10日 4 PM: 99.94%10月10日 5 PM: 100.00%10月10日 5 PM: 99.77%10月10日 6 PM: 99.76%10月10日 6 PM: 99.88%10月10日 7 PM: 100.00%10月10日 7 PM: 100.00%10月10日 8 PM: 99.93%10月10日 8 PM: 99.93%10月10日 9 PM: 100.00%10月10日 9 PM: 100.00%10月10日 10 PM: 99.93%10月10日 10 PM: 99.93%10月10日 11 PM: 99.93%10月10日 11 PM: 100.00%10月11日 8 AM: 99.36%10月11日 9 AM: 100.00%10月11日 9 AM: 99.36%10月11日 10 AM: 100.00%10月11日 10 AM: 99.32%10月11日 11 AM: 100.00%10月11日 11 AM: 100.00%10月11日 12 PM: 99.49%10月11日 12 PM: 100.00%10月11日 1 PM: 100.00%10月11日 1 PM: 100.00%10月11日 2 PM: 100.00%10月11日 2 PM: 99.11%10月11日 3 PM: 100.00%10月11日 3 PM: 100.00%10月11日 4 PM: 100.00%10月11日 4 PM: 100.00%10月11日 5 PM: 99.93%10月11日 5 PM: 99.93%10月11日 6 PM: 100.00%10月11日 6 PM: 99.85%10月11日 7 PM: 100.00%10月11日 7 PM: 99.93%10月11日 8 PM: 100.00%10月11日 8 PM: 100.00%10月11日 9 PM: 100.00%10月11日 9 PM: 100.00%10月11日 10 PM: 99.85%10月11日 10 PM: 99.92%10月11日 11 PM: 100.00%10月11日 11 PM: 100.00%

Reasoning

Toggleable

Supported Response Formats

OpenAI Chat CompletionsAnthropic MessagesOpenAI Responses

Total Context

512,000

Max Output

128,000

Input Price

¥2.1 / 1M tokens

Output Price

¥8.4 / 1M tokens

Cache Write

¥2.625 / 1M tokens

Cache Read

¥0.42 / 1M tokens

t
tokenmall-cdtx
国内

TTFT

6.85s

Throughput

55.51tps

Uptime

100.00%

Provider Model

tokenmall-cdtx/minimax/MiniMax-M3

Supported Parameters

temperaturetop_ptop_k

Recent Uptime

10月11日 11 PM100.00%
10月9日 7 PM: 99.90%10月9日 7 PM: 99.61%10月9日 8 PM: 100.00%10月9日 8 PM: 99.75%10月9日 9 PM: 100.00%10月9日 9 PM: 100.00%10月9日 10 PM: 100.00%10月9日 10 PM: 99.52%10月9日 11 PM: 100.00%10月9日 11 PM: 100.00%10月10日 8 AM: 100.00%10月10日 9 AM: 100.00%10月10日 9 AM: 100.00%10月10日 10 AM: 99.86%10月10日 10 AM: 100.00%10月10日 11 AM: 99.75%10月10日 11 AM: 99.88%10月10日 12 PM: 99.88%10月10日 12 PM: 99.79%10月10日 1 PM: 97.44%10月10日 1 PM: 100.00%10月10日 2 PM: 100.00%10月10日 2 PM: 98.96%10月10日 3 PM: 100.00%10月10日 3 PM: 99.91%10月10日 4 PM: 100.00%10月10日 4 PM: 100.00%10月10日 5 PM: 100.00%10月10日 5 PM: 99.08%10月10日 6 PM: 100.00%10月10日 6 PM: 100.00%10月10日 7 PM: 100.00%10月10日 7 PM: 99.93%10月10日 8 PM: 99.86%10月10日 8 PM: 100.00%10月10日 9 PM: 100.00%10月10日 9 PM: 100.00%10月10日 10 PM: 100.00%10月10日 10 PM: 100.00%10月10日 11 PM: 100.00%10月10日 11 PM: 99.93%10月11日 8 AM: 100.00%10月11日 9 AM: 100.00%10月11日 9 AM: 100.00%10月11日 10 AM: 100.00%10月11日 10 AM: 100.00%10月11日 11 AM: 100.00%10月11日 11 AM: 100.00%10月11日 12 PM: 100.00%10月11日 12 PM: 100.00%10月11日 1 PM: 100.00%10月11日 1 PM: 100.00%10月11日 2 PM: 100.00%10月11日 2 PM: 100.00%10月11日 3 PM: 100.00%10月11日 3 PM: 100.00%10月11日 4 PM: 100.00%10月11日 4 PM: 99.93%10月11日 5 PM: 100.00%10月11日 5 PM: 97.50%10月11日 6 PM: 100.00%10月11日 6 PM: 100.00%10月11日 7 PM: 100.00%10月11日 7 PM: 100.00%10月11日 8 PM: 99.93%10月11日 8 PM: 99.93%10月11日 9 PM: 100.00%10月11日 9 PM: 100.00%10月11日 10 PM: 100.00%10月11日 10 PM: 100.00%10月11日 11 PM: 100.00%10月11日 11 PM: 100.00%

Reasoning

Toggleable

Supported Response Formats

OpenAI Chat Completions

Total Context

512,000

Max Output

128,000

Input Price-40%

¥1.26 / 1M tokens

Output Price-40%

¥5.04 / 1M tokens

Cache Write-40%

¥1.575 / 1M tokens

Cache Read-40%

¥0.126 / 1M tokens

Performance for minimax/MiniMax-M3

Compare different providers across Zhinao API

Throughput

57.62 tok/s

TTFT

2.09 s

Uptime for minimax/MiniMax-M3

Uptime for minimax/MiniMax-M3 across all providers

Sample code and API for minimax/MiniMax-M3

Get API Key

Zhinao API normalizes requests and responses across providers for you

View full docs
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.360.cn/v1",
  apiKey: process.env.ZHINAO_API_KEY,
});

const response = await client.chat.completions.create({
  model: "minimax/MiniMax-M3",
  messages: [
    { role: "user", content: "Hello, how are you?" }
  ],
  temperature: 0.7,
  max_tokens: 1000,
});

console.log(response.choices[0].message.content);