The latest Gemini 3.7 Flash, DeepSeek-V4-Pro and Grok 4.6 models are now supported.

AI model pricing & performance

Compare every OpenAI, Gemini, Grok, and DeepSeek model supported by AI Text Summarizer, including price, relative intelligence, speed, and the jobs each model fits best.

Prices checked: August 14, 2026

Quick picks

A practical starting point before comparing every row.

Quality first

Grok 4.6

GPT-5.6 Sol

Choose for sustained high-quality reasoning, professional work, and demanding production tasks.

Best balance

Gemini 3.7 Flash

GPT-5.6 Luna

Strong capability at a lower cost, while keeping solid speed for everyday work.

Fast, high volume

Gemini 3.1 Flash-Lite

Gemini 3.5 Flash-Lite

A better fit for frequent summaries where responsiveness matters most.

Lowest cloud price

GPT-5 nano

Gemini 2.5 Flash-Lite

The two lowest list prices among cloud models on this page for simple, high-volume work.

Model price, intelligence & speed

Compare standard API input and output prices, relative intelligence, speed, and recommended use in one table.

GPT-5.6 SolLatest
gpt-5.6-sol
Input
$5.00
Output
$30.00

Intelligence

5/5

Speed

1/5

Complex reasoning and professional work

GPT-5.6 TerraLatest
gpt-5.6-terra
Input
$2.00
Output
$12.00

Intelligence

5/5

Speed

3/5

Complex reasoning and professional work

GPT-5.6 LunaLatest
gpt-5.6-luna
Input
$0.20
Output
$1.20

Intelligence

4/5

Speed

4/5

Everyday summaries and balanced workloads

GPT-5.5
gpt-5.5-2026-04-23
Input
$5.00
Output
$30.00

Intelligence

4/5

Speed

2/5

Complex reasoning and professional work

GPT-5.4
gpt-5.4-2026-03-05
Input
$2.50
Output
$15.00

Intelligence

4/5

Speed

3/5

Everyday summaries and balanced workloads

GPT-5.4 mini
gpt-5.4-mini-2026-03-17
Input
$0.75
Output
$4.50

Intelligence

3/5

Speed

4/5

General-purpose work with low latency

GPT-5.4 nano
gpt-5.4-nano-2026-03-17
Input
$0.20
Output
$1.25

Intelligence

2/5

Speed

3/5

Fast, frequent, high-volume summaries

GPT-5
gpt-5-2025-08-07
Input
$1.25
Output
$10.00

Intelligence

3/5

Speed

2/5

Everyday summaries and balanced workloads

GPT-5 mini
gpt-5-mini-2025-08-07
Input
$0.25
Output
$2.00

Intelligence

2/5

Speed

2/5

Simple extraction and maximum cost efficiency

GPT-5 nano
gpt-5-nano-2025-08-07
Input
$0.05
Output
$0.40

Intelligence

1/5

Speed

4/5

Simple extraction and maximum cost efficiency

GPT-4.1
gpt-4.1-2025-04-14
Input
$2.00
Output
$8.00

Intelligence

1/5

Speed

2/5

General-purpose work with low latency

Gemini 3.7 FlashLatest
gemini-3.7-flash
Input
$0.75
Output
$3.75

Intelligence

5/5

Speed

5/5

Everyday summaries and balanced workloads

Gemini 3.6 Flash
gemini-3.6-flash
Input
$0.75
Output
$3.75

Intelligence

4/5

Speed

5/5

Everyday summaries and balanced workloads

Gemini 3.5 Flash
gemini-3.5-flash
Input
$1.50
Output
$9.00

Intelligence

4/5

Speed

4/5

Everyday summaries and balanced workloads

Gemini 3.5 Flash-Lite
gemini-3.5-flash-lite
Input
$0.30
Output
$2.50

Intelligence

3/5

Speed

4/5

Fast, frequent, high-volume summaries

Gemini 3.1 Flash-Lite
gemini-3.1-flash-lite
Input
$0.25
Output
$1.50

Intelligence

2/5

Speed

5/5

Fast, frequent, high-volume summaries

Gemini 3.1 Pro
gemini-3.1-pro-preview
Input
$2.00
Output
$12.00

Intelligence

4/5

Speed

3/5

Complex multimodal and agentic work

Gemini 3 Flash
gemini-3-flash-preview
Input
$0.50
Output
$3.00

Intelligence

3/5

Speed

4/5

General-purpose work with low latency

Gemini 2.5 Flash
gemini-2.5-flash
Input
$0.30
Output
$2.50

Intelligence

1/5

Speed

5/5

Simple extraction and maximum cost efficiency

Gemini 2.5 Flash-Lite
gemini-2.5-flash-lite
Input
$0.10
Output
$0.40

Intelligence

1/5

Speed

5/5

Simple extraction and maximum cost efficiency

Grok 4.6Latest
grok-4.6
Input
$2.00
Output
$6.00

Intelligence

5/5

Speed

1/5

Complex reasoning and professional work

Grok 4.5
grok-4.5
Input
$2.00
Output
$6.00

Intelligence

4/5

Speed

2/5

Complex reasoning and professional work

Grok 4.3
grok-4.3
Input
$1.25
Output
$2.50

Intelligence

3/5

Speed

3/5

Everyday summaries and balanced workloads

DeepSeek

API pricing
DeepSeek-V4-ProLatest
deepseek-v4-pro
Input
$0.43
Output
$0.87

Intelligence

4/5

Speed

2/5

Complex reasoning and professional work

DeepSeek-V4-Flash
deepseek-v4-flash
Input
$0.14
Output
$0.28

Intelligence

4/5

Speed

3/5

Everyday summaries and balanced workloads

How to read this comparison

  1. 01Every price is the standard paid-tier list price, in USD per 1M tokens. Tax, batch discounts, caching, and tool fees are not included.
  2. 02Gemini output prices count thinking tokens. Heavier reasoning therefore costs more even when the visible answer is the same length.
  3. 03Long prompts are billed at a higher rate: OpenAI above 272K input tokens, Gemini 3.1 Pro and Grok 4.6 / 4.5 / 4.3 from 200K. DeepSeek is listed at its cache-miss rate, and cache hits cost less.
  4. 04Intelligence is a 1–5 band over the median of three platforms: the Artificial Analysis Intelligence Index, LMArena Elo, and the llm-stats Reasoning Index. Speed bands Artificial Analysis and OpenRouter output tokens per second rather than first-token latency, which is why the strongest reasoning models land low.
  5. 05DeepSeek bills by time of day from 2026-08-16. Peak is 01:00–04:00 and 06:00–10:00 UTC, and off-peak costs half the peak rate.
  6. 06Gemini 3.6 and 3.7 Flash are on an introductory rate until 2026-12-31. From 2027-01-01 they bill at $1.50 input and $7.50 output.