DeepSeek R1 Distill Llama 70B
DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.
| Provider | Status | Input | Output | Limits | Uptime | Speed | Notes |
|---|---|---|---|---|---|---|---|
| Novita | available | $0.80/Mtok | $0.80/Mtok | 8K tokens context | 100.0% | 993 ms p50 TTFT | bf16 |