Skip to content
SOLVENCY

Head to head

Gemini 3.7 Flash vs GPT-4.1

Gemini 3.7 Flash: currentGPT-4.1: legacyheavy tier

No verdict

One cost is measured by its benchmark and the other is modelled by Solvency. The two bases are never ranked against each other, so no verdict is given here.

Source: Artificial Analysis (artificialanalysis.ai) · verified 2026-08-21 · prices from each provider's pricing page

Input $/M tokens

lower is better

  • Gemini 3.7 Flash$0.75
  • GPT-4.1$2

Output $/M tokens

lower is better

  • Gemini 3.7 Flash$3.75
  • GPT-4.1$8

Cost per solved task

different bases — not comparable

  • Gemini 3.7 Flash$2.12
  • GPT-4.1$5.15
Gemini 3.7 FlashGPT-4.1
Input $/Mtok$0.75$2
Output $/Mtok$3.75$8
Cached input $/Mtok$0.075$0.5
Pass rate60%52%
Pass rate sourceAAAider
Cost basismeasuredmodelled
Cost per solved task$2.12$5.15

verified 2026-08-21 · heavy task tier; measured rows ignore the tier· Methodology· Gemini 3.7 Flash· GPT-4.1