GPT-6.1-Sol OpenAI released GPT-6.1 Sol, a model the company says delivers near-Astra performance at a lower cost for complex coding, computer use, and professional work, priced at $2.00 per 1M input tokens and $10.00 per 1M output tokens. The model supports reasoning.effort levels of low, medium (default), high, xhigh, and max, but not none or minimal, and requires the Responses API for tool calling, with Chat Completions supported without tool calling. GPT-6.1 Sol offers US and EU data residency, though Fast mode is unavailable with EU data residency, and prompts exceeding 272K input tokens are billed at 2x input and cache rates and 1.5x output for the full request. Near-Astra performance for complex work at a lower cost. Near-Astra performance for complex work at a lower cost. CompareTry in Playground Reasoning Highest Speed Fast Price $2•$10 Input•Output Input Text, Image Output Text GPT-6.1 Sol delivers near-Astra performance at a lower cost for complex coding, computer use, and professional work. Compare it with Astra on your tasks to assess the tradeoff between quality and cost. reasoning.effort supports low, medium default , high, xhigh, and max. The none and minimal reasoning efforts are not supported. Use the Responses API for tool calling. Chat Completions is supported without tool calling. GPT-6.1 Sol supports US and EU data residency. Fast mode is unavailable with EU data residency. See data residency eligibility. Pricing is based on the number of tokens used, or other metrics based on the model type. For tool-specific models, like search and computer use, there’s a fee per tool call. See details in the pricing page. Text tokens Per 1M tokens Input $2.00 Cached input $0.10 Cache writes $2.50 Output $10.00 Cached input tokens are priced at 5% of the uncached input token rate. Cache writes are billed at 1.25x the uncached input token rate. Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request. Fast mode prices are 2x Standard. Batch and Flex prices are 50% lower than Standard. Regional processing adds a 10% premium where available. Modalities Text Input and output Image Input only Audio Not supported Video Not supported Endpoints Live v1/live/sessions Chat Completions v1/chat/completions Responses v1/responses Realtime v1/realtime Realtime translation v1/realtime/translations Realtime transcription v1/realtime/transcription sessions Assistants v1/assistants Batch v1/batch Fine-tuning v1/fine-tuning Embeddings v1/embeddings Image generation v1/images/generations Videos v1/videos Image edit v1/images/edits Speech generation v1/audio/speech Transcription v1/audio/transcriptions Translation v1/audio/translations Moderation v1/moderations Completions legacy v1/completions Features Streaming Supported Function calling Supported Structured outputs Supported Fine-tuning Not supported Tools Tools supported by this model when using the Responses API. Web search Supported File search Supported Image generation Supported Code interpreter Supported Hosted shell Supported Apply patch Supported Skills Supported Computer use Supported MCP Supported Tool search Supported Snapshots Use gpt-6.1-sol to select this model. gpt-6.1-sol gpt-6.1-sol gpt-6.1-sol Rate limits Rate limits ensure fair and reliable access to the API by placing specific caps on requests, tokens, audio duration, or other usage within a given time period. Your usage tier determines how high these limits are set and automatically increases as you send more requests and spend more on the API.