Routin
Loading...
Loading...
面向低时延、高并发与成本敏感场景,强调快速响应与灵活推理部署,支持四档位思考与多模态理解能力。
A quick read of the model identity, commercial setup, and access footprint.
Published rates, tier transitions, and modality-specific surcharges in the current catalog entry.
| Context Length | Prompt | Output | Cache hit |
|---|---|---|---|
| 0 - 32,000 tokens 0-32k | ¥0.1500/1M tokens | ¥1.5000/1M tokens | ¥0.0300/1M tokens |
| 32,000 - 128,000 tokens | ¥0.3000/1M tokens | ¥3.0000/1M tokens | ¥0.0600/1M tokens |
| >= 128,000 tokens 128lk-256k | ¥0.6000/1M tokens | ¥6.0000/1M tokens | ¥0.0300/1M tokens |
Hourly usage, latency, reliability, and throughput for this model.
Top 10 AI coding agents using this model during the last 15 days.
| # | AI Coding Agents Usage | Total Tokens | Requests | Success rate | Last active |
|---|---|---|---|---|---|
| #1 | OpenCoWork Observed client-provider traffic for this model | 3.21M | 1.2K | 97.0% | 2026-08-03 |
| #2 | OpenCoWork Observed client-provider traffic for this model | 3.21M | 1.2K | 97.0% | 2026-08-03 |