Head to head
Qwen3 32BvsLlama 3.1 70B Instruct
For tasks, about a month.
Verdict · measured basis
Llama 3.1 70B Instruct costs $0.0005 per solved task against
Qwen3 32B at $0.0016 — ▼ 3.0x cheaper.
Over 200 tasks that is $0.21 a month.
On output token price alone Qwen3 32B looks 1.4x cheaper. The ranking reverses once you divide by pass rate — the cheaper tokens do not win the task.
Verified pricing, pass rate and cost per solved task for Qwen3 32B and Llama 3.1 70B Instruct| Figure | Qwen3 32B | Llama 3.1 70B Instruct |
|---|
| Input $/Mtok | $0.08 | $0.4 |
|---|
| Output $/Mtok | $0.28 | $0.4 |
|---|
| Cached input $/Mtok | $0.08 | $0.4 |
|---|
| Pass rate | 83% | 50% |
|---|
| Cost per solved task | $0.0016measured | $0.0005measured |
|---|
| Cost per month | $0.32 | $0.11 |
|---|
Cost per month against tasks per month · drag the marker
Solid measured · dashed modeled · dotted stale. The gap is the verdict.
Solvency Bench (solvency.dev) — first-party measurement · verified 2026-08-26 · prices from each provider's pricing page; measured rows ignore the tier · Methodology · All models at these settings