Routin
Loading...
Loading...
OpenAI o3-mini-high is the same model as o3-mini with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. The model features three adjustable reasoning effort levels and supports key developer capabilities including function calling, structured outputs, and streaming, though it does not include vision processing capabilities. The model demonstrates significant improvements over its predecessor, with expert testers preferring its responses 56% of the time and noting a 39% reduction in major errors on complex questions. With medium reasoning effort settings, o3-mini matches the performance of the larger o1 model on challenging reasoning evaluations like AIME and GPQA, while maintaining lower latency and cost.
同系列或更低输入价的公开模型。 / Same provider or cheaper catalog alternatives.
| Model | Provider | Type | Input | Output |
|---|---|---|---|---|
| text-embedding-3-small | OpenAI | chat | 0.0200 Token / 1M tokens | — |
| gpt-5-nano | OpenAI | chat | 0.0500 Token / 1M tokens | 0.4000 Token / 1M tokens |
| gpt-5-nano-2025-08-07 | OpenAI | chat | 0.0500 Token / 1M tokens | 0.4000 Token / 1M tokens |
| gpt-4.1-nano | OpenAI | chat | 0.1000 Token / 1M tokens | 0.4000 Token / 1M tokens |
| gpt-4.1-nano-2025-04-14 | OpenAI | chat | 0.1000 Token / 1M tokens |
A quick read of the model identity, commercial setup, and access footprint.
Published rates, tier transitions, and modality-specific surcharges in the current catalog entry.
| Pricing | Billing mode |
|---|---|
| Prompt | $0.0786/1M tokens |
| Output | $0.3143/1M tokens |
| Cache hit | $0.0393/1M tokens |
Hourly usage, latency, reliability, and throughput for this model.
Top 10 AI coding agents using this model during the last 15 days.
Top 10 official SDKs and language HTTP clients using this model during the last 15 days.
| 0.4000 Token / 1M tokens |
| sora-2 | OpenAI | video | 0.1000 Token / 1M tokens | — |