GPT-6 Sol Model | OpenAI API

OpenAI Developers

2 min read Original article ↗

gpt-6-sol

Built to power complex coding and agentic workflows.

Built to power complex coding and agentic workflows.

GPT-6 Sol is built for complex coding and agentic workflows. See GPT-6.1 Sol for the newer Sol model.

reasoning.effort supports none, low, medium (default), high, xhigh, and max. Use the Responses API for built-in tools and function calling. Chat Completions supports function calling only with reasoning_effort set to none.

EU data residency is available with Standard, Flex, and Batch processing. See data residency eligibility.

128,000 max output tokens

Apr 20, 2026 knowledge cutoff

Pricing

Pricing is based on the number of tokens used, or other metrics based on the model type. For tool-specific models, like search and computer use, there’s a fee per tool call. See details in the pricing page.

Cached input tokens are priced at 10% of the uncached input token rate.

Cache writes are billed at 1.25x the uncached input token rate.

Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request.

Regional processing adds a 10% premium where available. EU data residency is available with Standard, Flex, and Batch processing.

Batch and Flex are priced at 50% of Standard rates. Fast mode is priced at 2x the applicable rates.

Endpoints

Chat Completions

v1/chat/completions

Realtime translation

v1/realtime/translations

Realtime transcription

v1/realtime/transcription_sessions

Fine-tuning

v1/fine-tuning

Image generation

v1/images/generations

Image edit

v1/images/edits

Speech generation

v1/audio/speech

Transcription

v1/audio/transcriptions

Translation

v1/audio/translations

Completions (legacy)

v1/completions

Features

Function calling

Supported

Structured outputs

Supported

Tools

Tools supported by this model when using the Responses API.

Image generation

Supported

Code interpreter

Supported

Snapshots

Use gpt-6-sol in your API requests.

gpt-6-sol

Rate limits

Rate limits ensure fair and reliable access to the API by placing specific caps on requests, tokens, audio duration, or other usage within a given time period. Your usage tier determines how high these limits are set and automatically increases as you send more requests and spend more on the API.

TierRPMTPMBatch queue limit
FreeNot supported
Tier 1500500,0001,500,000
Tier 25,0001,000,0003,000,000
Tier 35,0002,000,000100,000,000
Tier 410,0004,000,000200,000,000
Tier 515,00040,000,00015,000,000,000