Deepinfra via Requesty
Meta Llama 3.1 70B Instruct
A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most.
- Context
- 130.8K
- Max output
- —
- Input / 1M
- $0.23
- Output / 1M
- $0.40
- Output speed
- —
Accepts
- Text input
Produces
- Text output