AI model pricing & performance

Compare every OpenAI and Gemini model supported by AI Text Summarizer, including price, relative intelligence, speed, and the jobs each model fits best.

Prices checked: July 31, 2026

Quick picks

A practical starting point before comparing every row.

Quality first

GPT-5.6 Sol / Gemini 3.5 Flash

Choose for sustained high-quality reasoning, professional work, and demanding production tasks.

Best balance

GPT-5.6 Terra / Gemini 3.6 Flash

Strong intelligence without paying flagship-model prices on every request.

Fast, high volume

GPT-5.6 Luna / Gemini 3.5 Flash-Lite

A better fit for frequent summaries where responsiveness matters most.

Lowest cloud price

GPT-5 nano

The lowest listed paid API price among the cloud models on this page.

Model price, intelligence & speed

Compare standard API input and output prices, relative intelligence, speed, and recommended use in one table.

OpenAI

GPT-5.6 Sol
gpt-5.6-sol
Input
$5.00
Output
$30.00
Intelligence
5/5

Frontier

Speed
2/5

Moderate

Complex reasoning and professional work

GPT-5.6 Terra
gpt-5.6-terra
Input
$2.00
Output
$12.00
Intelligence
4/5

Very high

Speed
4/5

Very fast

Everyday summaries and balanced workloads

GPT-5.6 Luna
gpt-5.6-luna
Input
$0.20
Output
$1.20
Intelligence
3/5

High

Speed
5/5

Fastest

Fast, frequent, high-volume summaries

GPT-5.5
gpt-5.5-2026-04-23
Input
$5.00
Output
$30.00
Intelligence
4/5

Very high

Speed
2/5

Moderate

Complex reasoning and professional work

GPT-5.4
gpt-5.4-2026-03-05
Input
$2.50
Output
$15.00
Intelligence
4/5

Very high

Speed
2/5

Moderate

Complex reasoning and professional work

GPT-5.4 mini
gpt-5.4-mini-2026-03-17
Input
$0.75
Output
$4.50
Intelligence
3/5

High

Speed
4/5

Very fast

Everyday summaries and balanced workloads

GPT-5.4 nano
gpt-5.4-nano-2026-03-17
Input
$0.20
Output
$1.25
Intelligence
2/5

Capable

Speed
5/5

Fastest

Simple extraction and maximum cost efficiency

GPT-5
gpt-5-2025-08-07
Input
$1.25
Output
$10.00
Intelligence
4/5

Very high

Speed
2/5

Moderate

Complex reasoning and professional work

GPT-5 mini
gpt-5-mini-2025-08-07
Input
$0.25
Output
$2.00
Intelligence
3/5

High

Speed
4/5

Very fast

Everyday summaries and balanced workloads

GPT-5 nano
gpt-5-nano-2025-08-07
Input
$0.05
Output
$0.40
Intelligence
2/5

Capable

Speed
5/5

Fastest

Simple extraction and maximum cost efficiency

GPT-4.1
gpt-4.1-2025-04-14
Input
$2.00
Output
$8.00
Intelligence
3/5

High

Speed
4/5

Very fast

General-purpose work with low latency

Google Gemini

Gemini 3.6 Flash
gemini-3.6-flash
Input
$1.50
Output
$7.50
Intelligence
4/5

Very high

Speed
4/5

Very fast

Everyday summaries and balanced workloads

Gemini 3.5 Flash
gemini-3.5-flash
Input
$1.50
Output
$9.00
Intelligence
5/5

Frontier

Speed
4/5

Very fast

Complex reasoning and professional work

Gemini 3.5 Flash-Lite
gemini-3.5-flash-lite
Input
$0.30
Output
$2.50
Intelligence
3/5

High

Speed
5/5

Fastest

Fast, frequent, high-volume summaries

Gemini 3.1 Flash-Lite
gemini-3.1-flash-lite
Input
$0.25
Output
$1.50
Intelligence
3/5

High

Speed
5/5

Fastest

Fast, frequent, high-volume summaries

Gemini 3.1 ProPreview
gemini-3.1-pro-preview
Input
$2.00
Output
$12.00
Intelligence
5/5

Frontier

Speed
2/5

Moderate

Complex multimodal and agentic work

Gemini 3 FlashPreview
gemini-3-flash-preview
Input
$0.50
Output
$3.00
Intelligence
4/5

Very high

Speed
4/5

Very fast

General-purpose work with low latency

Gemini 2.5 Flash
gemini-2.5-flash
Input
$0.30
Output
$2.50
Intelligence
3/5

High

Speed
4/5

Very fast

General-purpose work with low latency

Gemini 2.5 Flash-Lite
gemini-2.5-flash-lite
Input
$0.10
Output
$0.40
Intelligence
2/5

Capable

Speed
5/5

Fastest

Simple extraction and maximum cost efficiency

How to read this comparison

  1. 01Prices are standard paid-tier list prices in USD per 1 million tokens. Taxes, batch discounts, caching, tool calls, and provider-specific charges are not included.
  2. 02Gemini output pricing includes thinking tokens, so reasoning can affect the final output-token charge.
  3. 03OpenAI charges higher long-context rates above 272K input tokens. Gemini 3.1 Pro uses higher rates for prompts above 200K tokens.
  4. 04Intelligence and speed are editorial relative guides derived from each provider’s product positioning. Real latency changes with prompt length, reasoning, region, and service load.