Compare models

Qwen2.5 72B Instruct vs Mixtral 8x7B Instruct vs Llama 3.3 70B Instruct

Every figure below carries the source that reported it and the moment it was observed. Values that failed validation are withheld rather than shown, and a dash means no source we track carries that field for that model.
Qwen2.5 72B Instruct vs Mixtral 8x7B Instruct vs Llama 3.3 70B Instruct, compared across 14 tracked fields. Each value shows the source that reported it and when.
FieldQwen2.5 72B InstructMixtral 8x7B InstructLlama 3.3 70B Instruct
Downloads, 30 days
293,852Synced from sourceCurrent4h ago
207,778Synced from sourceCurrent4h ago
904,760 (highest of the compared values)Synced from sourceCurrent4h ago
Likes
997Synced from sourceCurrent4h ago
4,757 (highest of the compared values)Synced from sourceCurrent4h ago
3,066Synced from sourceCurrent4h ago
Context window
32,768Synced from sourceCurrent4h ago
—
131,072 (highest of the compared values)Synced from sourceCurrent4h ago
Input, $/M tokens
0.36Synced from sourceCurrent4h ago
—
0.1 (lowest of the compared values)Synced from sourceCurrent4h ago
Modality
textSeeded, unreviewedwritten 1mo ago
textSeeded, unreviewedwritten 1mo ago
textSeeded, unreviewedwritten 1mo ago
Licence
qwenSeeded, unreviewedwritten 1mo ago
apache-2.0Seeded, unreviewedwritten 1mo ago
llama-3.3Seeded, unreviewedwritten 1mo ago
Gated
noSynced from sourceCurrent4h ago
noSynced from sourceCurrent4h ago
manualSynced from sourceCurrent4h ago
Last updated
2025-01-12Synced from sourceCurrent4h ago
2025-07-24Synced from sourceCurrent4h ago
2024-12-21Synced from sourceCurrent4h ago
Library
transformersSynced from sourceCurrent4h ago
vllmSynced from sourceCurrent4h ago
transformersSynced from sourceCurrent4h ago
Task
text-generationSynced from sourceCurrent4h ago
—
text-generationSynced from sourceCurrent4h ago
Hugging Face
Qwen/Qwen2.5-72B-InstructSynced from sourceCurrent4h ago
mistralai/Mixtral-8x7B-Instruct-v0.1Synced from sourceCurrent4h ago
meta-llama/Llama-3.3-70B-InstructSynced from sourceCurrent4h ago
OpenRouter id
qwen/qwen-2.5-72b-instructSynced from sourceCurrent4h ago
—
meta-llama/llama-3.3-70b-instructSynced from sourceCurrent4h ago
Output, $/M tokens
0.4Synced from sourceCurrent4h ago
—
0.32Synced from sourceCurrent4h ago
Vendor
AlibabaSeeded, unreviewedwritten 1mo ago
Mistral AISeeded, unreviewedwritten 1mo ago
MetaSeeded, unreviewedwritten 1mo ago

Sources: Hugging Face Hub, OpenRouter model catalogue. Most recent observation 4h ago. Where a row is marked, ▲ is the highest of the values shown and ▼ the lowest — arithmetic on the figures above, not a ranking or a recommendation. Rows where neither extreme is meaningful are left unmarked rather than given a direction they do not have.

Change the comparison

Remove one, or open the directory to pick a different set. Up to 3 at a time.


Every figure on this site resolves to the source that reported it, the moment it was observed, and how it was measured. A value that fails validation is withheld rather than shown — withheld means we refused to publish it, not that it is zero.