# GPT-6.1-Sol

> Source: <https://developers.openai.com/api/docs/models/gpt-6.1-sol>
> Published: 2026-09-29 17:07:19+00:00

Near-Astra performance for complex work at a lower cost.

Near-Astra performance for complex work at a lower cost.

CompareTry in Playground

Reasoning

Highest

Speed

Fast

Price

$2•$10

Input•Output

Input

Text, Image

Output

Text

GPT-6.1 Sol delivers near-Astra performance at a lower cost for complex coding,
computer use, and professional work. Compare it with Astra on your tasks to
assess the tradeoff between quality and cost.

reasoning.effort supports low, medium (default), high, xhigh, and
max. The none and minimal reasoning efforts are not supported.

Use the Responses API for tool calling. Chat Completions is supported without
tool calling.

GPT-6.1 Sol supports US and EU data residency. Fast mode is unavailable with EU
data residency. See data residency eligibility.

Pricing is based on the number of tokens used, or other metrics based on the model type. For tool-specific models, like search and computer use, there’s a fee per tool call. See details in the pricing page.

Text tokens

Per 1M tokens

Input

$2.00

Cached input

$0.10

Cache writes

$2.50

Output

$10.00

Cached input tokens are priced at 5% of the uncached input token rate.

Cache writes are billed at 1.25x the uncached input token rate.

Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request.

Fast mode prices are 2x Standard. Batch and Flex prices are 50% lower than Standard.

Regional processing adds a 10% premium where available.

Modalities

Text

Input and output

Image

Input only

Audio

Not supported

Video

Not supported

Endpoints

Live

v1/live/sessions

Chat Completions

v1/chat/completions

Responses

v1/responses

Realtime

v1/realtime

Realtime translation

v1/realtime/translations

Realtime transcription

v1/realtime/transcription_sessions

Assistants

v1/assistants

Batch

v1/batch

Fine-tuning

v1/fine-tuning

Embeddings

v1/embeddings

Image generation

v1/images/generations

Videos

v1/videos

Image edit

v1/images/edits

Speech generation

v1/audio/speech

Transcription

v1/audio/transcriptions

Translation

v1/audio/translations

Moderation

v1/moderations

Completions (legacy)

v1/completions

Features

Streaming

Supported

Function calling

Supported

Structured outputs

Supported

Fine-tuning

Not supported

Tools

Tools supported by this model when using the Responses API.

Web search

Supported

File search

Supported

Image generation

Supported

Code interpreter

Supported

Hosted shell

Supported

Apply patch

Supported

Skills

Supported

Computer use

Supported

MCP

Supported

Tool search

Supported

Snapshots

Use gpt-6.1-sol to select this model.

gpt-6.1-sol

gpt-6.1-sol

gpt-6.1-sol

Rate limits

Rate limits ensure fair and reliable access to the API by placing specific caps on requests, tokens, audio duration, or other usage within a given time period. Your usage tier determines how high these limits are set and automatically increases as you send more requests and spend more on the API.
