Routin
Loading...
Loading...
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning and coding performance across benchmarks like AIME (99.5% with Python) and SWE-bench, outperforming its predecessor o3-mini and even approaching o3 in some domains. Despite its smaller size, o4-mini exhibits high accuracy in STEM tasks, visual problem solving (e.g., MathVista, MMMU), and code editing. It is especially well-suited for high-throughput scenarios where latency or cost is critical. Thanks to its efficient architecture and refined reinforcement learning training, o4-mini can chain tools, generate structured outputs, and solve multi-step tasks with minimal delay—often in under a minute.
同系列或更低输入价的公开模型。 / Same provider or cheaper catalog alternatives.
| Model | Provider | Type | Input | Output |
|---|---|---|---|---|
| text-embedding-3-small | OpenAI | chat | 0.0200 Token / 1M tokens | — |
| gpt-5-nano | OpenAI | chat | 0.0500 Token / 1M tokens | 0.4000 Token / 1M tokens |
| gpt-5-nano-2025-08-07 | OpenAI | chat | 0.0500 Token / 1M tokens | 0.4000 Token / 1M tokens |
| gpt-4.1-nano | OpenAI | chat | 0.1000 Token / 1M tokens | 0.4000 Token / 1M tokens |
| gpt-4.1-nano-2025-04-14 | OpenAI | chat | 0.1000 Token / 1M tokens |
A quick read of the model identity, commercial setup, and access footprint.
Published rates, tier transitions, and modality-specific surcharges in the current catalog entry.
| Pricing | Billing mode |
|---|---|
| Prompt | $0.0786/1M tokens |
| Output | $0.3143/1M tokens |
| Cache hit | $0.0196/1M tokens |
Hourly usage, latency, reliability, and throughput for this model.
Top 10 AI coding agents using this model during the last 15 days.
Top 10 official SDKs and language HTTP clients using this model during the last 15 days.
| 0.4000 Token / 1M tokens |
| sora-2 | OpenAI | video | 0.1000 Token / 1M tokens | — |
| # | SDK / Language | Total Tokens | Requests | Success rate | Last active |
|---|---|---|---|---|---|
| #1 | Node.js Language SDK / HTTP client | 165.2K | 64 | 100.0% | 2026-09-16 |
| #2 | Python Language SDK / HTTP client | 49.4K | 848 | 99.1% | 2026-09-17 |
| #3 | Go Language SDK / HTTP client |
| 15.8K |
| 133 |
| 98.5% |
| 2026-09-17 |
| #4 | OpenAI SDK Official OpenAI client libraries | 341 | 1 | 100.0% | 2026-09-07 |
| #5 | JavaScript Language SDK / HTTP client | 305 | 1 | 100.0% | 2026-09-09 |