Processing...Please wait while we secure this action
DeepSeek model

R1 Distill Llama 70B

DeepSeek: R1 Distill Llama 70B is a DeepSeek model listed in the NextModel catalog for Chinese Q&A, general chat workloads. Its listed price is $0.152 / 1M tokens input and $0.152 / 1M tokens output per 1M tokens, with a 8.2k token context window.

DeepSeekPlatform curatedCatalog
Streaming
Input price$0.152 / 1M tokens
Output price$0.152 / 1M tokens
Context length8.2k tokens
Max output8.2k tokens

What is R1 Distill Llama 70B in NextModel?

DeepSeek: R1 Distill Llama 70B is a DeepSeek model listed in the NextModel catalog for Chinese Q&A, general chat workloads. Its listed price is $0.152 / 1M tokens input and $0.152 / 1M tokens output per 1M tokens, with a 8.2k token context window.

Best use cases

  • Chinese Q&A
  • general chat

OpenAI-compatible code example

Keep the OpenAI SDK style, set base_url to NextModel, and use the catalog model ID deepseek--deepseek-r1-distill-llama-70b.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="deepseek/deepseek-r1-distill-llama-70b",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Similar alternatives

DeepSeekCatalog

Available through the NextModel gateway via DeepSeek.

$0 / 1M tokensInput$0 / 1M tokensOutputContext
Best forChinese Q&A, general chat
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
DeepSeekCatalog

Available through the NextModel gateway via DeepSeek.

$0 / 1M tokensInput$0 / 1M tokensOutputContext
Best forChinese Q&A, general chat
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
DeepSeekCatalog

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...

$0.038 / 1M tokensInput$0.152 / 1M tokensOutput163.8kContext
Best forChinese Q&A, general chat
RoutingConfigured
StreamingTool callingJSON modeLong context
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

DeepSeek: R1 Distill Llama 70B API questions