GPT 5.6 Sol 20% price reduction

85 points by izakfr 17 hours ago on hackernews | 73 comments

Models

gpt-5.6-sol

Frontier model for complex professional work

Frontier model for complex professional work

GPT-5.6 Sol is the frontier model in the GPT-5.6 family. It roughly corresponds to the unsuffixed model tier used in earlier GPT-5 families. The gpt-5.6 alias routes requests to GPT-5.6 Sol. Reasoning.effort supports: none, low, medium (default), high, xhigh, and max.

128,000

max output tokens

Feb 16, 2026 knowledge cutoff

Pricing

Pricing is based on the number of tokens used, or other metrics based on the model type. For tool-specific models, like search and computer use, there’s a fee per tool call. See details in the

pricing page.

GPT-5.6 Sol costs $4 per million input tokens and $20 per million output tokens, a 20% reduction in input pricing and a 33% reduction in output pricing. GPT-5.6 Sol’s promotional pricing is available at least through November 21, 2026.

Prompts with >272K input tokens are priced at 2x input and 1.5x output for the full request.

Cache writes are billed at 1.25x the uncached input token rate.

Endpoints

Chat Completions

v1/chat/completions

Realtime translation

v1/realtime/translations

Realtime transcription

v1/realtime/transcription_sessions

Fine-tuning

v1/fine-tuning

Image generation

v1/images/generations

Image edit

v1/images/edits

Speech generation

v1/audio/speech

Transcription

v1/audio/transcriptions

Translation

v1/audio/translations

Completions (legacy)

v1/completions

Features

Function calling

Supported

Structured outputs

Supported

Tools

Tools supported by this model when using the Responses API.

Image generation

Supported

Code interpreter

Supported

Snapshots

Snapshots let you lock in a specific version of the model so that performance and behavior remain consistent. Below is a list of all available snapshots and aliases for

GPT-5.6 Sol

.

gpt-5.6-sol

Rate limits

Rate limits ensure fair and reliable access to the API by placing specific caps on requests, tokens, audio duration, or other usage within a given time period. Your usage tier determines how high these limits are set and automatically increases as you send more requests and spend more on the API.

TierRPMTPMBatch queue limit
FreeNot supported
Tier 1500500,0001,500,000
Tier 25,0001,000,0003,000,000
Tier 35,0002,000,000100,000,000
Tier 410,0004,000,000200,000,000
Tier 515,00040,000,00015,000,000,000