Skip to content
SOLVENCY

Head to head

Gemini 3.1 Pro (preview) vs Grok 4.6

Gemini 3.1 Pro (preview): currentGrok 4.6: currentheavy tier

No verdict

Cost per solved task needs a published pass rate for both models. At least one has none, so it is reported as missing rather than estimated.

Source: Scale SEAL leaderboard, SWE-bench Pro · verified 2026-08-21 · prices from each provider's pricing page

Input $/M tokens

lower is better

  • Gemini 3.1 Pro (preview)$2
  • Grok 4.6$2

Output $/M tokens

lower is better

  • Gemini 3.1 Pro (preview)$12
  • Grok 4.6$6

Cost per solved task

needs a pass rate for both

  • Gemini 3.1 Pro (preview)$4.30
  • Grok 4.6missing
Gemini 3.1 Pro (preview)Grok 4.6
Input $/Mtok$2$2
Output $/Mtok$12$6
Cached input $/Mtok$0.2$0.5
Pass rate46%missing
Pass rate sourceSEALmissing
Cost basismodelledmissing
Cost per solved task$4.30missing

verified 2026-08-21 · heavy task tier; measured rows ignore the tier· Methodology· Gemini 3.1 Pro (preview)· Grok 4.6