Skip to contentx402 is live on Robinhood Chain: agents can now pay for any model, per call, in USDG →
OmniRail
← All models

DeepSeek · Language model

R1 Distill Llama 70B

deepseek/deepseek-r1-distill-llama-70b
  • Language
Released Jan 23, 2025

DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...

Specifications

Context window
8,192 tokens
Max output
7,372 tokens
Input
text
Output
text
Reasoning
supported
Tokenizer
Llama3
Knowledge cutoff
2024-07-31

Supported parameters

Accepted by the model and forwarded as-is by the gateway.

  • frequency_penalty
  • max_tokens
  • presence_penalty
  • reasoning
  • repetition_penalty
  • seed
  • stop
  • temperature
  • top_k
  • top_p

Use it

POST /api/v1/chat/completions · MCP chat · full reference

curl https://omnirail.org/api/v1/chat/completions \
  -H "Authorization: Bearer $OMNIRAIL_KEY" -H "Content-Type: application/json" \
  -d '{"model":"deepseek/deepseek-r1-distill-llama-70b","messages":[{"role":"user","content":"Hello"}],"max_tokens":300}'