Deepinfra via Requesty
Meta Llama 3.1 405B Instruct
A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most.
- Context
- 130.8K
- Max output
- —
- Input / 1M
- $0.80
- Output / 1M
- $0.80
- Output speed
- —
Accepts
- Text input
Produces
- Text output