Compare models

Llama 3.1 8B Instruct vs DeepSeek R1 vs Llama 3.3 70B Instruct

Every figure below carries the source that reported it and the moment it was observed. Values that failed validation are withheld rather than shown, and a dash means no source we track carries that field for that model.
Llama 3.1 8B Instruct vs DeepSeek R1 vs Llama 3.3 70B Instruct, compared across 14 tracked fields. Each value shows the source that reported it and when.
FieldLlama 3.1 8B InstructDeepSeek R1Llama 3.3 70B Instruct
Downloads, 30 days
6,225,344 (highest of the compared values)Synced from sourceCurrent2h ago
815,126Synced from sourceCurrent2h ago
917,910Synced from sourceCurrent2h ago
Likes
7,861Synced from sourceCurrent2h ago
14,285 (highest of the compared values)Synced from sourceCurrent2h ago
3,058Synced from sourceCurrent2h ago
Context window
131,072 (highest of the compared values)Synced from sourceCurrent2h ago
64,000Synced from sourceCurrent2h ago
131,072 (highest of the compared values)Synced from sourceCurrent2h ago
Input, $/M tokens
0.05 (lowest of the compared values)Synced from sourceCurrent2h ago
0.7Synced from sourceCurrent2h ago
0.1Synced from sourceCurrent2h ago
Modality
textSeeded, unreviewedwritten 29d ago
reasoningSeeded, unreviewedwritten 29d ago
textSeeded, unreviewedwritten 29d ago
Licence
llama-3.1Seeded, unreviewedwritten 29d ago
mitSeeded, unreviewedwritten 29d ago
llama-3.3Seeded, unreviewedwritten 29d ago
Gated
manualSynced from sourceCurrent2h ago
noSynced from sourceCurrent2h ago
manualSynced from sourceCurrent2h ago
Last updated
2024-09-25Synced from sourceCurrent2h ago
2025-03-27Synced from sourceCurrent2h ago
2024-12-21Synced from sourceCurrent2h ago
Library
transformersSynced from sourceCurrent2h ago
transformersSynced from sourceCurrent2h ago
transformersSynced from sourceCurrent2h ago
Task
text-generationSynced from sourceCurrent2h ago
text-generationSynced from sourceCurrent2h ago
text-generationSynced from sourceCurrent2h ago
Hugging Face
meta-llama/Llama-3.1-8B-InstructSynced from sourceCurrent2h ago
deepseek-ai/DeepSeek-R1Synced from sourceCurrent2h ago
meta-llama/Llama-3.3-70B-InstructSynced from sourceCurrent2h ago
OpenRouter id
meta-llama/llama-3.1-8b-instructSynced from sourceCurrent2h ago
deepseek/deepseek-r1Synced from sourceCurrent2h ago
meta-llama/llama-3.3-70b-instructSynced from sourceCurrent2h ago
Output, $/M tokens
0.08Synced from sourceCurrent2h ago
2.5Synced from sourceCurrent2h ago
0.32Synced from sourceCurrent2h ago
Vendor
MetaSeeded, unreviewedwritten 29d ago
DeepSeekSeeded, unreviewedwritten 29d ago
MetaSeeded, unreviewedwritten 29d ago

Sources: Hugging Face Hub, OpenRouter model catalogue. Most recent observation 2h ago. Where a row is marked, ▲ is the highest of the values shown and ▼ the lowest — arithmetic on the figures above, not a ranking or a recommendation. Rows where neither extreme is meaningful are left unmarked rather than given a direction they do not have.

Change the comparison

Remove one, or open the directory to pick a different set. Up to 3 at a time.


Every figure on this site resolves to the source that reported it, the moment it was observed, and how it was measured. A value that fails validation is withheld rather than shown — withheld means we refused to publish it, not that it is zero.