Compare models

GPT-4o vs Llama 3.3 70B Instruct

Every figure below carries the source that reported it and the moment it was observed. Values that failed validation are withheld rather than shown, and a dash means no source we track carries that field for that model.
GPT-4o vs Llama 3.3 70B Instruct, compared across 14 tracked fields. Each value shows the source that reported it and when.
FieldGPT-4oLlama 3.3 70B Instruct
Downloads, 30 days—
904,760Synced from sourceCurrent1h ago
Likes—
3,066Synced from sourceCurrent1h ago
Context window
128,000Synced from sourceCurrent1h ago
131,072 (highest of the compared values)Synced from sourceCurrent1h ago
Input, $/M tokens
2.5Synced from sourceCurrent1h ago
0.1 (lowest of the compared values)Synced from sourceCurrent1h ago
Modality
textSeeded, unreviewedwritten 1mo ago
textSeeded, unreviewedwritten 1mo ago
Licence
proprietarySeeded, unreviewedwritten 1mo ago
llama-3.3Seeded, unreviewedwritten 1mo ago
Gated—
manualSynced from sourceCurrent1h ago
Last updated—
2024-12-21Synced from sourceCurrent1h ago
Library—
transformersSynced from sourceCurrent1h ago
Task—
text-generationSynced from sourceCurrent1h ago
Hugging Face—
meta-llama/Llama-3.3-70B-InstructSynced from sourceCurrent1h ago
OpenRouter id
openai/gpt-4oSynced from sourceCurrent1h ago
meta-llama/llama-3.3-70b-instructSynced from sourceCurrent1h ago
Output, $/M tokens
10Synced from sourceCurrent1h ago
0.32Synced from sourceCurrent1h ago
Vendor
OpenAISeeded, unreviewedwritten 1mo ago
MetaSeeded, unreviewedwritten 1mo ago

Sources: Hugging Face Hub, OpenRouter model catalogue. Most recent observation 1h ago. Where a row is marked, ▲ is the highest of the values shown and ▼ the lowest — arithmetic on the figures above, not a ranking or a recommendation. Rows where neither extreme is meaningful are left unmarked rather than given a direction they do not have.

Change the comparison

Remove one, or open the directory to pick a different set. Up to 3 at a time.


Every figure on this site resolves to the source that reported it, the moment it was observed, and how it was measured. A value that fails validation is withheld rather than shown — withheld means we refused to publish it, not that it is zero.