![]()
Built to power complex coding and agentic workflows.
Built to power complex coding and agentic workflows.
GPT-6 Sol is built for complex coding and agentic workflows. See GPT-6.1 Sol for the newer Sol model.
reasoning.effort supports none, low, medium (default), high, xhigh, and max.
Use the Responses API for built-in tools and function calling. Chat Completions supports function calling only with reasoning_effort set to none.
EU data residency is available with Standard, Flex, and Batch processing. See data residency eligibility.
128,000 max output tokens
Apr 20, 2026 knowledge cutoff
Pricing
Pricing is based on the number of tokens used, or other metrics based on the model type. For tool-specific models, like search and computer use, there’s a fee per tool call. See details in the pricing page.
Cached input tokens are priced at 10% of the uncached input token rate.
Cache writes are billed at 1.25x the uncached input token rate.
Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request.
Regional processing adds a 10% premium where available. EU data residency is available with Standard, Flex, and Batch processing.
Batch and Flex are priced at 50% of Standard rates. Fast mode is priced at 2x the applicable rates.
Endpoints
Chat Completions
v1/chat/completions
Realtime translation
v1/realtime/translations
Realtime transcription
v1/realtime/transcription_sessions
Fine-tuning
v1/fine-tuning
Image generation
v1/images/generations
Image edit
v1/images/edits
Speech generation
v1/audio/speech
Transcription
v1/audio/transcriptions
Translation
v1/audio/translations
Completions (legacy)
v1/completions
Features
Function calling
Supported
Structured outputs
Supported
Tools
Tools supported by this model when using the Responses API.
Image generation
Supported
Code interpreter
Supported
Snapshots
Use gpt-6-sol in your API requests.
![]()
Rate limits
Rate limits ensure fair and reliable access to the API by placing specific caps on requests, tokens, audio duration, or other usage within a given time period. Your usage tier determines how high these limits are set and automatically increases as you send more requests and spend more on the API.
| Tier | RPM | TPM | Batch queue limit |
|---|---|---|---|
| Free | Not supported | ||
| Tier 1 | 500 | 500,000 | 1,500,000 |
| Tier 2 | 5,000 | 1,000,000 | 3,000,000 |
| Tier 3 | 5,000 | 2,000,000 | 100,000,000 |
| Tier 4 | 10,000 | 4,000,000 | 200,000,000 |
| Tier 5 | 15,000 | 40,000,000 | 15,000,000,000 |