OpenAI API: models and pricing

FastMetal serves 18 OpenAI models through one OpenAI-compatible API key, billed per token from a prepaid balance. Change the base URL and key, and existing code calls them as is. The list also shows OpenAI models we do not serve yet.

Sort by:

56 models

GPT-6.1 Sol

gpt-6.1-sol

GPT-6.1 Sol is an upgrade to GPT-6 Sol from OpenAI, positioned below the flagship GPT-6 Astra in the GPT-6 series. It is suited for agentic coding, computer use, document-heavy professional...

Provider:openai logoopenai
Pricing:$2.11 in · $10.55 out / 1M tokens
Context:1.1M tokens

GPT-6.1 Sol Pro

gpt-6.1-sol-pro

**Cost note:** pro mode spends far more...

Provider:openai logoopenai
Pricing:$2.20 in · $11.00 out / 1M tokens
Context:1.1M tokens

GPT-6 Luna

gpt-6-luna
Policy unknown

GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...

Provider:openai logoopenai
Pricing:$0.1055 in · $0.5275 out / 1M tokens
Context:1.1M tokens

GPT-6 Luna Pro

gpt-6-luna-pro

Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

Provider:openai logoopenai
Pricing:$0.11 in · $0.55 out / 1M tokens
Context:1.1M tokens

GPT-6 Sol

gpt-6-sol
Policy unknown

GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...

Provider:openai logoopenai
Pricing:$2.11 in · $10.55 out / 1M tokens
Context:1.1M tokens

GPT-6 Sol Pro

gpt-6-sol-pro

Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

Provider:openai logoopenai
Pricing:$2.20 in · $11.00 out / 1M tokens
Context:1.1M tokens

GPT-6 Astra

gpt-6-astra
Policy unknown

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

Provider:openai logoopenai
Pricing:$10.55 in · $52.75 out / 1M tokens
Context:1.1M tokens

GPT-6 Astra Pro

gpt-6-astra-pro

Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

Provider:openai logoopenai
Pricing:$11.00 in · $55.00 out / 1M tokens
Context:1.1M tokens

GPT-5.6 Luna

gpt-5.6-luna

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

Provider:openai logoopenai
Pricing:$0.211 in · $1.27 out / 1M tokens
Context:1.1M tokens

GPT-5.6 Sol

gpt-5.6-sol

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

Provider:openai logoopenai
Pricing:$2.20 in · $11.00 out / 1M tokens
Context:1.1M tokens

GPT-5.6 Terra

gpt-5.6-terra

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

Provider:openai logoopenai
Pricing:$2.11 in · $12.66 out / 1M tokens
Context:1.1M tokens

GPT-5.5

openai/gpt-5.5

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:1.1M tokens
Text Overall
#30

GPT-5.5 Pro

openai/gpt-5.5-pro

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:1.1M tokens

GPT-5.4 Image 2

openai/gpt-5.4-image-2

It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:272K tokens

GPT-5.4 Mini

openai/gpt-5.4-mini

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens

GPT-5.4 Nano

gpt-5.4-nano
Policy unknown

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks.

Provider:openai logoopenai
Pricing:$0.211 in · $1.32 out / 1M tokens
Context:400K tokens

GPT-5.4

openai/gpt-5.4

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for text and image inputs, enabling high-context reasoning, codi…

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:1.1M tokens
Text Overall
#53

GPT-5.4 Pro

openai/gpt-5.4-pro

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:1.1M tokens

GPT-5.3 Chat

openai/gpt-5.3-chat

GPT-5.3 Chat is an update to ChatGPT's most-used model that makes everyday conversations smoother, more useful, and more directly helpful.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens

GPT-5.3-Codex

openai/gpt-5.3-codex

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens

GPT-5.2-Codex

openai/gpt-5.2-codex

GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens

GPT-5.2

openai/gpt-5.2

GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens
Text Overall
#104

GPT-5.2 Chat

openai/gpt-5.2-chat

GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens

GPT-5.2 Pro

openai/gpt-5.2-pro

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens

GPT-5.1-Codex-Max

openai/gpt-5.1-codex-max

GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens

GPT-5.1

openai/gpt-5.1

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens
Text Overall
#97

GPT-5.1 Chat

openai/gpt-5.1-chat

GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens

GPT-5.1-Codex

openai/gpt-5.1-codex

GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens

GPT-5.1-Codex-Mini

openai/gpt-5.1-codex-mini

GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens

GPT-5

openai/gpt-5

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy in high-stakes us…

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens

GPT-5 Chat

openai/gpt-5-chat

GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens
Text Overall
#116

GPT-5 Mini

gpt-5-mini

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost.

Provider:openai logoopenai
Pricing:$0.2638 in · $2.11 out / 1M tokens
Context:400K tokens

GPT-5 Nano

openai/gpt-5-nano

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:400K tokens

gpt-oss-120b

gpt-oss-120b

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases.

Provider:openai logoopenai
Pricing:$0.102 in · $0.5035 out / 1M tokens
Context:131K tokens
Text Overall
#212

gpt-oss-120b (exacto)

openai/gpt-oss-120b:exacto

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:131K tokens

gpt-oss-120b (free)

openai/gpt-oss-120b:free

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:131K tokens

gpt-oss-20b

gpt-oss-20b

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for lower-latency inference and deplo…

Provider:openai logoopenai
Pricing:$0.0317 in · $0.1477 out / 1M tokens
Context:131K tokens
Text Overall
#263

gpt-oss-20b (free)

openai/gpt-oss-20b:free

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for lower-latency inference and deplo…

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:131K tokens

o3

openai/o3

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:200K tokens

o4 Mini

openai/o4-mini

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:200K tokens

GPT-4.1

openai/gpt-4.1

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:1.0M tokens

GPT-4.1 Mini

gpt-4.1-mini

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost.

Provider:openai logoopenai
Pricing:$0.44 in · $1.76 out / 1M tokens
Context:1.0M tokens

GPT-4.1 Nano

gpt-4.1-nano

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million token context window, and scores 80.1% on MMLU, 50.3% on GPQA, a…

Provider:openai logoopenai
Pricing:$0.11 in · $0.44 out / 1M tokens
Context:1.0M tokens

o3 Mini High

openai/o3-mini-high

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding…

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:200K tokens
Text Overall
#201

o3 Mini

openai/o3-mini

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:200K tokens
Text Overall
#218

o1

openai/o1

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason using chain of thought.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:200K tokens

GPT-4o (2024-08-06)

openai/gpt-4o-2024-08-06

The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/).

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens
Text Overall
#238

GPT-4o-mini

gpt-4o-mini

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs.

Provider:openai logoopenai
Pricing:$0.1583 in · $0.633 out / 1M tokens
Context:128K tokens

GPT-4o-mini (2024-07-18)

openai/gpt-4o-mini-2024-07-18

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens
Text Overall
#262

GPT-4o

openai/gpt-4o

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as fast and 50% more cost-effec…

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens

GPT-4o (2024-05-13)

openai/gpt-4o-2024-05-13

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as fast and 50% more cost-effec…

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens
Text Overall
#222

GPT-4 Turbo

openai/gpt-4-turbo

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens

GPT-4 Turbo (older v1106)

openai/gpt-4-1106-preview

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to April 2023.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:128K tokens
Text Overall
#270

GPT-3.5 Turbo

openai/gpt-3.5-turbo

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:16K tokens

GPT-4

openai/gpt-4

OpenAI's flagship model, GPT-4 is a large-scale multimodal language model capable of solving difficult problems with greater accuracy than previous models due to its broader general knowledge and advanced reasoning capabilities.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:8K tokens

GPT-4 (older v0314)

openai/gpt-4-0314

GPT-4-0314 is the first version of GPT-4 released, with a context length of 8,192 tokens, and was supported until June 14. Training data: up to Sep 2021.

Provider:openai logoopenai
Pricing:Not on FastMetal yet
Context:8K tokens
Text Overall
#293