Routin
Loading...
Loading...
Hy3 正式版面向真实业务场景打磨,采用 295B 总参数、21B 激活 MoE 架构,原生支持 256K 上下文,并提供no_think(极速响应)、think_low(快速思考)与 think_high(深度推理) 多档思考模式,兼顾极速响应、复杂推理与调用成本。相比 Preview 版本,Hy3 基于腾讯元宝、WorkBuddy、ima 、Marvis等真实业务反馈,重点提升了 Coding Agent、长文理解、多轮上下文承接、搜索问答与复杂任务执行能力,在减少幻觉、提升任务完成度和工程可用性方面表现更稳。更加适合前端任务、跨文件代码开发、长文档分析、办公自动化和多步骤 Agent 工作流等实用场景。
同系列或更低输入价的公开模型。 / Same provider or cheaper catalog alternatives.
| Model | Provider | Type | Input | Output |
|---|---|---|---|---|
| MiniMax-M2.1 | MiniMax | chat | 2.73 Token / 1M tokens | 10.92 Token / 1M tokens |
| MiniMax-M2.1-highspeed | MiniMax | chat | 5.46 Token / 1M tokens | 21.84 Token / 1M tokens |
| MiniMax-M2.5 | MiniMax | chat | 0.2730 Token / 1M tokens | 10.92 Token / 1M tokens |
| MiniMax-M2.5-highspeed | MiniMax | chat | 5.46 Token / 1M tokens | 21.84 Token / 1M tokens |
| MiniMax-M2.7 | MiniMax | chat | 2.73 Token / 1M tokens | 10.92 Token / 1M tokens |
A quick read of the model identity, commercial setup, and access footprint.
Published rates, tier transitions, and modality-specific surcharges in the current catalog entry.
| Pricing | Billing mode |
|---|---|
| Prompt | $0.1114/1M tokens |
| Output | $0.4457/1M tokens |
| Cache hit | $0.0279/1M tokens |
Hourly usage, latency, reliability, and throughput for this model.
Top 10 AI coding agents using this model during the last 15 days.
Top 10 official SDKs and language HTTP clients using this model during the last 15 days.
| MiniMax-M2.7-highspeed | MiniMax | chat | 5.46 Token / 1M tokens | 16.80 Token / 1M tokens |
| # | SDK / Language | Total Tokens | Requests | Success rate | Last active |
|---|---|---|---|---|---|
| #1 | Python Language SDK / HTTP client | 0 | 1 | 0.0% | 2026-08-20 |