All models

Magnum v4 72B vs GPT-3.5 Turbo 16k

Magnum v4 72B at $3.45 in and $5.75 out and GPT-3.5 Turbo 16k at $3.45 in and $4.60 out — per million tokens, from the same wallet.

A

Magnum v4 72B

Anthracite

2.2¢

for a client proposal

$3.45 in · $5.75 out / 1M

16K context · released October 2024

Full page →

GPT-3.5 Turbo 16k

OpenAI

2.1¢

for a client proposal

$3.45 in · $4.60 out / 1M

16K context · released August 2023

Full page →

Asking all both at once — identical prompt, identical context — costs about 4.3¢ for a typical client proposal.

At a glance

  • All 2 cost the same on input.
  • Cheapest output: GPT-3.5 Turbo 16k at $4.60 per million tokens.
  • Largest context: GPT-3.5 Turbo 16k at 16K tokens — about 23 pages.

Cost per task

Estimated totals at the rates above. Solid bar is input, lighter bar is output.

Quick code fix

2K in · 500 out
Magnum v4 72B
0.98¢
GPT-3.5 Turbo 16k
0.92¢

Draft a client proposal from notes

5K in · 800 out
Magnum v4 72B
2.2¢
GPT-3.5 Turbo 16k
2.1¢

Context-heavy session (Alyph workspace)

155K in · 4K out
Magnum v4 72B
needs 155K ctx
GPT-3.5 Turbo 16k
needs 155K ctx

Analyze a 50-page legal contract

35K in · 1K out
Magnum v4 72B
needs 35K ctx
GPT-3.5 Turbo 16k
needs 35K ctx

Large architecture refactor

250K in · 3K out
Magnum v4 72B
needs 250K ctx
GPT-3.5 Turbo 16k
needs 250K ctx

Solid part of each bar is input tokens, the lighter part is output. Models that don’t fit a task are marked instead of priced.

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$1.96$3.910 tokens250k500k750k1000k
Magnum v4 72B
GPT-3.5 Turbo 16k

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

Pricing

Magnum v4 72B
GPT-3.5 Turbo 16k
Input / 1M tokens
$3.45
$3.45
Output / 1M tokens
$5.75
$4.60

Specs

Magnum v4 72B
GPT-3.5 Turbo 16k
Released
October 2024
August 2023
Context window
16K
16K
Max output
2K
4K
Input
Text
Text
Output
Text
Text
Reasoning
Answers directly
Answers directly
Knowledge cutoff
2024-06-30
2021-09-30

Try it with real numbers

2K tokens in · 500 tokens out.

Magnum v4 72B

Anthracite

0.98¢

for this task

71% input · 29% output

Full pricing →

GPT-3.5 Turbo 16k

OpenAI

0.92¢

for this task

75% input · 25% output

Full pricing →

All 2, same prompt and context: 1.9¢. That is the entire cost of the comparison.

Your $5 welcome credit covers about 263 of these.

Which should you pick?

For long documents, GPT-3.5 Turbo 16k has the largest context window here — 16K tokens.

On price, GPT-3.5 Turbo 16k is the cheapest of the two on a typical task, at about 2.1¢.

Or don’t pick. Send the identical prompt to all both, read the answers side by side, and keep the winner. How to run a fair bake-off →

Questions, answered

Which is cheaper, Magnum v4 72B or GPT-3.5 Turbo 16k?

On a typical client proposal (5K tokens in, 800 out): GPT-3.5 Turbo 16k at 2.1¢; Magnum v4 72B at 2.2¢. Long prompts can change the order when long-context rates apply.

Which has the largest context window?

GPT-3.5 Turbo 16k — 16K tokens, about 23 pages.

Can I run Magnum v4 72B and GPT-3.5 Turbo 16k side by side on Alyph?

Yes — that is what Alyph is built for: the identical prompt with identical context to all 2, answers rendered side by side, billed from one wallet. This exact combination costs about 4.3¢ for a typical task.

Ask both at once

Identical prompt, identical context, answers side by side — about 4.3¢ for a typical client proposal. Free to start, with $5 of credit.