Skip to content
SOLVENCY

Head to head

Claude Opus 5 vs GPT-4.1

Claude Opus 5: currentGPT-4.1: legacyheavy tier

No verdict

One cost is measured by its benchmark and the other is modelled by Solvency. The two bases are never ranked against each other, so no verdict is given here.

Source: Artificial Analysis (artificialanalysis.ai) · verified 2026-08-21 · prices from each provider's pricing page

Input $/M tokens

lower is better

  • Claude Opus 5$5
  • GPT-4.1$2

Output $/M tokens

lower is better

  • Claude Opus 5$25
  • GPT-4.1$8

Cost per solved task

different bases — not comparable

  • Claude Opus 5$12.01
  • GPT-4.1$5.15
Claude Opus 5GPT-4.1
Input $/Mtok$5$2
Output $/Mtok$25$8
Cached input $/Mtok$0.5$0.5
Pass rate68%52%
Pass rate sourceAAAider
Cost basismeasuredmodelled
Cost per solved task$12.01$5.15

verified 2026-08-21 · heavy task tier; measured rows ignore the tier· Methodology· Claude Opus 5· GPT-4.1