chat.deepseek_r1_distill_llama_70b
chat.deepseek_r1_distill_llama_70b

Loading model information...

Input
DeepSeek R1 Distill Llama 70B
max_tokens4096
temperature1
top_p1
presence_penalty0
frequency_penalty0
DeepSeek R1 Distill Llama 70B Online
Hi! I'm a helpful AI assistant. What can I do for you?

DeepSeek R1 Distill Llama 70B API —— R1 级推理,没有 671B 的账单

DeepSeek R1 Distill Llama 70B API 把 DeepSeek-R1 的思维链蒸馏进稠密的 Llama 3.3 70B——当 8B 太轻、完整 671B 又超出所需时的折中之道。

在小蒸馏与 6710 亿参数旗舰之间,坐着那个明智的默认项。DeepSeek R1 Distill Llama 70B API 是把 DeepSeek-R1 推理蒸馏进稠密 Llama 3.3 70B 的模型:它保留让 R1 值得用的逐步思维链,比完整模型更快、便宜得多,又在真实的数学与逻辑上明显强过 8B 蒸馏。对多数推理工作来说,DeepSeek R1 Distill Llama 70B API 就是那个平衡点。

你通过 VibeToken、以 OpenAI chat-completions 协议调用 DeepSeek R1 Distill Llama 70B API;把模型设为 deepseek-r1-distill-llama-70b,客户端就绪。DeepSeek R1 Distill Llama 70B API 按输入输出统一每百万 token ¥0.80 计费,比标价低 5%,与覆盖整个目录的是同一把密钥。

折中之道

DeepSeek R1 Distill Llama 70B API 为何是明智默认

足够做真实推理的深度,价格又能一直开着。

均衡的蒸馏版

比 8B 更强,比 671B 更便宜更快——DeepSeek R1 Distill Llama 70B API 正是多数推理负载真正想要的折中之道。

R1 思维链

蒸馏自 DeepSeek-R1,DeepSeek R1 Distill Llama 70B API 逐步思考并流式输出轨迹,让你读到推理,而不只是结果。

稠密 Llama 3.3 70B

稠密 70B 主干带来稳定延迟与强数学逻辑,且无需调度专家混合。

每百万统一 ¥0.80

输入输出同价——DeepSeek R1 Distill Llama 70B API 按单一可预期费率计费,低于标价 5%,规模上易于预测。

开源权重

DeepSeek-R1-Distill-Llama-70B 为开源权重;你可自行评测,再让 DeepSeek R1 Distill Llama 70B API 托管运行,无需自跑 GPU。

OpenAI 形态

一个 chat-completions 端点、一把 VibeToken 密钥——指向 vibetoken.cn/v1、写上模型名,跳过 DeepSeek SDK。

RouterBase dashboard preview
接上线

三步完成对 DeepSeek R1 Distill Llama 70B API 的首次调用

从密钥到思维链,只需几分钟。

  1. 创建密钥

    一把 VibeToken 密钥即可访问 DeepSeek R1 Distill Llama 70B API 及目录里的每个模型。

  2. 指向并命名

    把任意 OpenAI 客户端发往 vibetoken.cn/v1,model=deepseek-r1-distill-llama-70b。流式与推理轨迹默认开启。

  3. 读取轨迹

    思维链与答案并排抵达;用量随每个响应返回,成本一目了然。

定价

按使用量付费

VibeToken 透传合作方价格,与模型官方公开 API 价格对比。

生产之选

把 DeepSeek R1 Distill Llama 70B API 设为默认的团队

规模化运行的平衡点。

Camila RojasLead Engineer, Fathom

我们把 DeepSeek R1 Distill Llama 70B API 设为默认——它推理得像 R1,却没有 R1 的账单。

Idris BelloCTO, Slate

8B 对我们的数学太轻,671B 又贵到不能常开。DeepSeek R1 Distill Llama 70B API 正好击中中间。

Hana KimML Lead, Everline

稠密 70B 意味着可预期的延迟;DeepSeek R1 Distill Llama 70B API 从不给我们的 SLO 添惊吓。

Viktor NovakStaff Engineer, Groundwork

输入输出统一 ¥0.80,预测变得毫不费力——月底不用再算进出比。

Renuka IyerFounder, Pathwise

DeepSeek R1 Distill Llama 70B API 流式输出的思维链好到我们原样保留。

Otto LindgrenPrincipal Engineer, Beacon Grid

开源权重让我们拿它和 671B 对标,然后 DeepSeek R1 Distill Llama 70B API 让我们上线更便宜的那个。

Zoe AlmeidaHead of AI, Rill

一把 VibeToken 密钥,我们只改一个字段就能把 DeepSeek R1 Distill Llama 70B API 和 Sonnet 做 A/B。

Jamal CarterBackend Lead, Overstory

我们把日常推理迁到 DeepSeek R1 Distill Llama 70B API,每任务成本下降,却没有一句质量抱怨。

Nina FalkEngineering Manager, Cindergrid

它现在是我们最先伸手去拿的推理模型——深度够,价格合理。

DeepSeek R1 Distill Llama 70B API —— 常见问题

把它设为默认前该权衡的。

VibeToken 对 DeepSeek-R1-Distill-Llama-70B 的托管接入——把 DeepSeek-R1 推理蒸馏进稠密 Llama 3.3 70B 的模型,通过兼容 OpenAI 的端点提供。