Deepinfra via Requesty
Llama 3.3 70B Instruct Turbo
A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most.
- Context
- 131.1K
- Max output
- —
- Input / 1M
- $0.12
- Output / 1M
- $0.30
- Output speed
- —
Accepts
- Text input
Produces
- Text output