Deepinfra via Requesty
Meta Llama 3.1 8B Instruct Turbo
A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most.
- Context
- 131.1K
- Max output
- —
- Input / 1M
- $0.02
- Output / 1M
- $0.05
- Output speed
- —
Accepts
- Text input
Produces
- Text output