Gemini 3.6 Flash vs Hermes 4 405B.

model vs model · prices as of

Gemini 3.6 Flash and Hermes 4 405B split the 5 rounds this table counts: the two list price rows, the cheapest price we can explain, the standard price and context. The weighting you pick decides the rest.

Google

Gemini 3.6 Flash

$3/M tokens

cheapest we can explain, at Kilo Gateway · $0.338 per 1,000 pages

1M context · tool calling · reads images

Nous Research

Hermes 4 405B

$2.10/M tokens

cheapest we can explain, at NanoGPT · $0.252 per 1,000 pages

131k context · open weights

NanoGPT

List price in (lower wins)

$0.75/M

$1/M

List price out (lower wins)

$3.75/M

$3/M

Cheapest we can explain (lower wins)

$3/M at Kilo Gateway

$2.10/M at NanoGPT

Standard price (lower wins)

$6/M

$6/M

Cost per 1,000 pages (lower wins)

$0.338

$0.252

Context

1M tokens

131k tokens

Open weights

no

yes

Tool calling

yes

no

Structured output

yes

yes

Reads images

yes

no

Reasoning

yes

yes

Released

21 Jul 2026

26 Aug 2025

Providers selling it

32

8

Price we measure against

the maker's own

OpenRouter

Trains on your prompts (no is better)

not stated

not stated

Providers selling both · 7

  • NanoGPT$6 / $2.10
  • Kilo Gateway$3 / $6
  • Eden AI$6 / $6
  • OpenRouter$6 / $6
  • Cortecs$6.21 / $6.19
  • TrustedRouter$6.33 / $6.33
  • Requesty$11.50 / $6

Gemini 3.6 Flash then Hermes 4 405B, per million tokens

Go further

Add a third or fourth model, or switch the weighting, in the tool.

What moved in the last seven days

  1. Gemini 3.6 FlashNous Portal$0.60 / $3 $0.75 / $3.7525% dearer ·
  2. Gemini 3.6 FlashOfox$1.50 / $7.50 $0.75 / $3.7550% cheaper ·

Frequently asked

Which is cheaper, Gemini 3.6 Flash or Hermes 4 405B?
Hermes 4 405B, at $2.10 per million tokens against $3 for Gemini 3.6 Flash, three parts input to one part output, read on 16 Sept 2026.
Where is each one cheapest?
Gemini 3.6 Flash: $3 at Kilo Gateway. Hermes 4 405B: $2.10 at NanoGPT.
Can I buy both from one provider?
Yes. 7 providers sell both: NanoGPT, Kilo Gateway, Eden AI, OpenRouter, Cortecs, TrustedRouter, and more. The cheapest of them for the two together is NanoGPT.
Which holds more context?
Gemini 3.6 Flash, at 1M tokens against 131k.

Prices exclude tax and are per million tokens, read 16 Sept 2026 from each provider's own catalog. The standard price blends three parts input to one part output; a page is 600 tokens in and 60 out. Bars are scaled to the larger value in each row.

Related head-to-headsGemini 3.6 Flash vs Gemini 3.7 FlashGemini 3.6 Flash vs Gemini 3.8 FlashGemini 3.6 Flash vs Gemini Flash LatestGemini 3.7 Flash vs Hermes 4 405BGemini 3.8 Flash vs Hermes 4 405BMorecompare any two modelsevery model we track