Llama 3.1 8B Instruct vs Qwen3.7 Flash.
model vs model · prices as of
Qwen3.7 Flash takes 3 of 5 rounds against Llama 3.1 8B Instruct. Llama 3.1 8B Instruct counters on price: $0.10 per million tokens against $0.14, for 131k context instead of 1M.
Llama 3.1 8B Instruct
$0.10/M tokens
cheapest we can explain, at Kilo Gateway · $0.0144 per 1,000 pages
131k context · tool calling · open weights
Qwen3.7 Flash
$0.14/M tokens
cheapest we can explain, at Eden AI · $0.0168 per 1,000 pages
1M context · tool calling · reads images
List price in (lower wins)
$0.05/M
$0.03/M
List price out (lower wins)
$0.08/M
$0.13/M
Cheapest we can explain (lower wins)
$0.10/M at Kilo Gateway
$0.14/M at Eden AI
Standard price (lower wins)
$0.23/M
$0.22/M
Cost per 1,000 pages (lower wins)
$0.0144
$0.0168
Context
131k tokens
1M tokens
Open weights
yes
not stated
Tool calling
yes
yes
Structured output
yes
yes
Reads images
no
yes
Reasoning
no
yes
Released
23 Jul 2024
27 Jul 2026
Providers selling it
24
15
Price we measure against
OpenRouter
the maker's own
Trains on your prompts (no is better)
not stated
not stated
Providers selling both · 6
Kilo Gateway$0.10 / $0.22
Nous Portal$0.10 / $0.22
TrustedRouter$0.12 / $0.23
OpenRouter$0.23 / $0.22
NanoGPT$0.25 / $0.22
Vercel AI Gateway$0.88 / $0.22
Llama 3.1 8B Instruct then Qwen3.7 Flash, per million tokens
Go further
Add a third or fourth model, or switch the weighting, in the tool.
What moved in the last seven days
- Llama 3.1 8B Instruct
Nous Portal$0.02 / $0.03 $0.02 / $0.0425% dearer · - Qwen3.7 Flash
Nous Portal$0.02 / $0.10 $0.03 / $0.1325% dearer ·
Frequently asked
- Which is cheaper, Llama 3.1 8B Instruct or Qwen3.7 Flash?
- Llama 3.1 8B Instruct, at $0.10 per million tokens against $0.14 for Qwen3.7 Flash, three parts input to one part output, read on 16 Sept 2026.
- Where is each one cheapest?
- Llama 3.1 8B Instruct: $0.10 at Kilo Gateway. Qwen3.7 Flash: $0.14 at Eden AI.
- Can I buy both from one provider?
- Yes. 6 providers sell both: Kilo Gateway, Nous Portal, TrustedRouter, OpenRouter, NanoGPT, Vercel AI Gateway. The cheapest of them for the two together is Kilo Gateway.
- Which holds more context?
- Qwen3.7 Flash, at 1M tokens against 131k.
Prices exclude tax and are per million tokens, read 16 Sept 2026 from each provider's own catalog. The standard price blends three parts input to one part output; a page is 600 tokens in and 60 out. Bars are scaled to the larger value in each row.