All models

Hermes 4 70B vs GPT-5.6 Luna

Hermes 4 70B at $0.149 in and $0.460 out and GPT-5.6 Luna at $0.115 in and $0.690 out — per million tokens, from the same wallet.

N

Hermes 4 70B

Nous Research

0.11¢

for a client proposal

$0.149 in · $0.460 out / 1M

131K context · released August 2025

Full page →

GPT-5.6 Luna

OpenAI

0.11¢

for a client proposal

$0.115 in · $0.690 out / 1M

1.05M context · released 1 month ago

Full page →

Asking all both at once — identical prompt, identical context — costs about 0.22¢ for a typical client proposal.

At a glance

  • Cheapest input: GPT-5.6 Luna at $0.115 per million tokens.
  • Cheapest output: Hermes 4 70B at $0.460 per million tokens.
  • Largest context: GPT-5.6 Luna at 1.05M tokens — about 1,500 pages.

Cost per task

Estimated totals at the rates above. Solid bar is input, lighter bar is output.

Quick code fix

2K in · 500 out
Hermes 4 70B
0.05¢
GPT-5.6 Luna
0.06¢

Draft a client proposal from notes

5K in · 800 out
Hermes 4 70B
0.11¢
GPT-5.6 Luna
0.11¢

Context-heavy session (Alyph workspace)

155K in · 4K out
Hermes 4 70B
needs 155K ctx
GPT-5.6 Luna
2.1¢

Analyze a 50-page legal contract

35K in · 1K out
Hermes 4 70B
0.57¢
GPT-5.6 Luna
0.47¢

Large architecture refactor

250K in · 3K out
Hermes 4 70B
needs 250K ctx
GPT-5.6 Luna
3.1¢

Solid part of each bar is input tokens, the lighter part is output. Models that don’t fit a task are marked instead of priced.

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$0.20$0.390 tokens250k500k750k1000k
Hermes 4 70B
GPT-5.6 Luna

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

Pricing

Hermes 4 70B
GPT-5.6 Luna
Input / 1M tokens
$0.149
$0.115
Output / 1M tokens
$0.460
$0.690
Cached input / 1M
$0.011
Web search
0.57¢ / search

Specs

Hermes 4 70B
GPT-5.6 Luna
Released
August 2025
1 month ago
Context window
131K
1.05M
Max output
128K
Input
Text
Files, Images, and Text
Output
Text
Text
Reasoning
Shows its thinking
Shows its thinking
Knowledge cutoff
2024-08-31
2026-02-16

Try it with real numbers

2K tokens in · 500 tokens out.

Hermes 4 70B

Nous Research

0.05¢

for this task

57% input · 43% output

Full pricing →

GPT-5.6 Luna

OpenAI

0.06¢

for this task

40% input · 60% output

Full pricing →

All 2, same prompt and context: 0.11¢. That is the entire cost of the comparison.

Your $5 welcome credit covers about 4,528 of these.

Which should you pick?

For long documents, GPT-5.6 Luna has the largest context window here — 1.05M tokens.

On price, Hermes 4 70B is the cheapest of the two on a typical task, at about 0.11¢.

For images or files, GPT-5.6 Luna can read them directly; Hermes 4 70B is text-only.

Or don’t pick. Send the identical prompt to all both, read the answers side by side, and keep the winner. How to run a fair bake-off →

Questions, answered

Which is cheaper, Hermes 4 70B or GPT-5.6 Luna?

On a typical client proposal (5K tokens in, 800 out): Hermes 4 70B at 0.11¢; GPT-5.6 Luna at 0.11¢. Long prompts can change the order when long-context rates apply.

Which has the largest context window?

GPT-5.6 Luna — 1.05M tokens, about 1,500 pages.

Can I run Hermes 4 70B and GPT-5.6 Luna side by side on Alyph?

Yes — that is what Alyph is built for: the identical prompt with identical context to all 2, answers rendered side by side, billed from one wallet. This exact combination costs about 0.22¢ for a typical task.

Ask both at once

Identical prompt, identical context, answers side by side — about 0.22¢ for a typical client proposal. Free to start, with $5 of credit.