Meta
Llama 3.1 405B
Meta's largest open-weights model, competitive with frontier closed models.
- Context
- 128K
- Max output
- 4.1K
- Input / 1M
- $3.50
- Output / 1M
- $3.50
- Output speed
- 30 t/s
Accepts
- Text input
Produces
- Text output
AI model comparison
Compare recorded price, context, output limits, speed, latency, capabilities, and input/output support. Every missing value stays visible, and unsourced benchmarks are excluded.
Llama 3.1 405B and Llama-3.3-70B-Instruct both come up when teams are choosing a model for production work, and the honest answer usually depends on your workload rather than a leaderboard. This page puts their published specifications side by side — context window, token pricing, supported inputs and outputs — so you can see where they actually differ.
Meta
Meta's largest open-weights model, competitive with frontier closed models.
Accepts
Produces
Meta
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
Accepts
Produces
At a glance
This is a directional summary of the values recorded in this directory—not a universal quality verdict. Test the finalists on your own prompts before committing.
Where Llama 3.1 405B leads in recorded data
No complete recorded factor currently favors this model. That does not establish lower real-world quality.
Where Llama-3.3-70B-Instruct leads in recorded data
Context capacity
Higher listed limit
Llama 3.1 405B:128K
Llama-3.3-70B-Instruct:128K
Same recorded value
Maximum output
Higher listed limit
Llama 3.1 405B:4.1K
Llama-3.3-70B-Instruct:4.1K
Same recorded value
Input price
Lower recorded price
Llama 3.1 405B:$3.50
Llama-3.3-70B-Instruct:$1.25
Llama-3.3-70B-Instruct
Output price
Lower recorded price
Llama 3.1 405B:$3.50
Llama-3.3-70B-Instruct:$1.25
Llama-3.3-70B-Instruct
Output speed
Higher recorded throughput
Llama 3.1 405B:30 t/s
Llama-3.3-70B-Instruct:—
No complete comparison
Time to first token
Lower recorded latency
Llama 3.1 405B:0.70s
Llama-3.3-70B-Instruct:—
No complete comparison
Intelligence Index
Higher directory index
Llama 3.1 405B:74.0
Llama-3.3-70B-Instruct:—
No complete comparison
Evidence
A score appears only when both records include a verified source. It is not treated as a universal ranking.
Specifications
| Specification | Llama 3.1 405B | Llama-3.3-70B-Instruct |
|---|---|---|
| Provider | Meta | Meta |
| Model family | Llama 3.1 | llama |
| Accepted inputs | text | text |
| Produced outputs | text | text |
| Context window | 128K | 128K |
| Max output | 4.1K | 4.1K |
| Intelligence Index | 74.0 | — |
| Input price / 1M | $3.50 | $1.25Recorded lead |
| Output price / 1M | $3.50 | $1.25Recorded lead |
| Output speed | 30 t/s | — |
| Time to first token | 0.70s | — |
| Reasoning mode | Not documented | Not documented |
| Tool calling | Documented | Documented |
| Image input | Not documented | Not documented |
| Audio input | Not documented | Not documented |
| Structured output | Documented | Not documented |
| License | Open weights | Open weights |
| Released | Jul 23, 2024 | Dec 6, 2024 |
Workload guidance
On the published numbers, Llama-3.3-70B-Instruct accepts the larger context window and Llama-3.3-70B-Instruct is cheaper per input token. Which matters more depends on whether your bottleneck is document size or spend — the table above has the exact figures.
Provenance
Prices and hosted performance can change. SyncDev records source references and clearly separates documented specifications from measured values; confirm commercial terms with the provider before deployment.
Llama 3.1 405B
Meta · model record and recorded source
Llama-3.3-70B-Instruct
Meta · model record and recorded source
Choose any two approved model records and review them with the same evidence rules.