All models

Mercury 2 vs Qwen2.5 VL 72B Instruct

Mercury 2 at $0.287 in and $0.863 out and Qwen2.5 VL 72B Instruct at $0.287 in and $0.863 out — per million tokens, from the same wallet.

I

Mercury 2

Inception

0.21¢

for a client proposal

$0.287 in · $0.863 out / 1M

128K context · released 5 months ago

Full page →
Q

Qwen2.5 VL 72B Instruct

Qwen

0.21¢

for a client proposal

$0.287 in · $0.863 out / 1M

32K context · released February 2025

Full page →

Asking all both at once — identical prompt, identical context — costs about 0.43¢ for a typical client proposal.

At a glance

  • All 2 cost the same on input.
  • Largest context: Mercury 2 at 128K tokens — about 180 pages.

Cost per task

Estimated totals at the rates above. Solid bar is input, lighter bar is output.

Quick code fix

2K in · 500 out
Mercury 2
0.1¢
Qwen2.5 VL 72B Instruct
0.1¢

Draft a client proposal from notes

5K in · 800 out
Mercury 2
0.21¢
Qwen2.5 VL 72B Instruct
0.21¢

Context-heavy session (Alyph workspace)

155K in · 4K out
Mercury 2
needs 155K ctx
Qwen2.5 VL 72B Instruct
needs 155K ctx

Analyze a 50-page legal contract

35K in · 1K out
Mercury 2
1.1¢
Qwen2.5 VL 72B Instruct
needs 35K ctx

Large architecture refactor

250K in · 3K out
Mercury 2
needs 250K ctx
Qwen2.5 VL 72B Instruct
needs 250K ctx

Solid part of each bar is input tokens, the lighter part is output. Models that don’t fit a task are marked instead of priced.

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$0.20$0.400 tokens250k500k750k1000k
Mercury 2
Qwen2.5 VL 72B Instruct

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

Pricing

Mercury 2
Qwen2.5 VL 72B Instruct
Input / 1M tokens
$0.287
$0.287
Output / 1M tokens
$0.863
$0.863
Cached input / 1M
$0.029

Specs

Mercury 2
Qwen2.5 VL 72B Instruct
Released
5 months ago
February 2025
Context window
128K
32K
Max output
50K
Input
Text
Text and Images
Output
Text
Text
Reasoning
Shows its thinking
Answers directly
Knowledge cutoff
2024-06-30

Try it with real numbers

2K tokens in · 500 tokens out.

Mercury 2

Inception

0.1¢

for this task

57% input · 43% output

Full pricing →

Qwen2.5 VL 72B Instruct

Qwen

0.1¢

for this task

57% input · 43% output

Full pricing →

All 2, same prompt and context: 0.2¢. That is the entire cost of the comparison.

Your $5 welcome credit covers about 2,484 of these.

Which should you pick?

For long documents, Mercury 2 has the largest context window here — 128K tokens.

On price they are identical: about 0.21¢ for a typical task each.

For images or files, Qwen2.5 VL 72B Instruct can read them directly; Mercury 2 is text-only.

If you want to watch the model think, Mercury 2 shows reasoning; the other answers directly.

Or don’t pick. Send the identical prompt to all both, read the answers side by side, and keep the winner. How to run a fair bake-off →

Questions, answered

Which is cheaper, Mercury 2 or Qwen2.5 VL 72B Instruct?

On a typical client proposal (5K tokens in, 800 out) they cost the same — about 0.21¢ each. Compare them on context and capabilities instead.

Which has the largest context window?

Mercury 2 — 128K tokens, about 180 pages.

Can I run Mercury 2 and Qwen2.5 VL 72B Instruct side by side on Alyph?

Yes — that is what Alyph is built for: the identical prompt with identical context to all 2, answers rendered side by side, billed from one wallet. This exact combination costs about 0.43¢ for a typical task.

Ask both at once

Identical prompt, identical context, answers side by side — about 0.43¢ for a typical client proposal. Free to start, with $5 of credit.