Deepinfra via Requesty
DeepSeek R1 Distill Llama 70B
DeepSeek-R1-Distill-Llama-70B is an open-weights, 70-billion parameter language model developed by DeepSeek. It distills the advanced mathematical, coding, and logical reasoning patterns of the flagship DeepSeek-R1 model into the highly optimized Llama-3.3-70B-Instruct architecture.
- Context
- 64K
- Max output
- 8.2K
- Input / 1M
- $0.23
- Output / 1M
- $0.69
- Output speed
- —
Accepts
- Text input
Produces
- Text output