GPT-6 Astra is OpenAI’s most capable model for complex reasoning, coding, computer use, research, and document creation, with strong performance on multistep workflows across code, browsers, and professional software. It supports reasoning.effort levels of low, medium, high, xhigh, and max, and achieves strong results while often using fewer output tokens for lower estimated cost per task than earlier models. It is also OpenAI’s most aligned model yet, designed to follow intent carefully, respect task boundaries, adapt to changing requirements, and communicate transparently.
Providers #
Route requests across multiple providers. Copy a provider slug to set your preference.
**$10-20**
*/ M tokens*
**$50-75**
*/ M tokens*
Read:
1-2/ M tokens
Write:
12.5-25/ M tokens1.05M1.02s39.7tps
**$10-20**
*/ M tokens*
**$50-75**
*/ M tokens*
Read:
1-2/ M tokens
Write:
12.5-25/ M tokens1.05M4.44s13.9tps
Uptime #
24hours Direct request success rate on AI Gateway and per-provider.
Throughput #
24hours P50 throughput on live AI Gateway traffic, in tokens per second (TPS).
Latency #
24hours P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds.
Activity #
Token volume and request traffic to this model over time.
Apps #
Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for. View All
Related Models #
More models from OpenAI