chat.deepseek_r1_distill_llama_70b
chat.deepseek_r1_distill_llama_70b

Loading model information...

Input
DeepSeek R1 Distill Llama 70B
max_tokens4096
temperature1
top_p1
presence_penalty0
frequency_penalty0
DeepSeek R1 Distill Llama 70B Online
Hi! I'm a helpful AI assistant. What can I do for you?

DeepSeek R1 Distill Llama 70B API —— R1 級推理,沒有 671B 的帳單

DeepSeek R1 Distill Llama 70B API 把 DeepSeek-R1 的思維鏈蒸餾進稠密的 Llama 3.3 70B——當 8B 太輕、完整 671B 又超出所需時的折中之道。

在小蒸餾與 6710 億參數旗艦之間,坐著那個明智的預設項。DeepSeek R1 Distill Llama 70B API 是把 DeepSeek-R1 推理蒸餾進稠密 Llama 3.3 70B 的模型:它保留讓 R1 值得用的逐步思維鏈,比完整模型更快、便宜得多,又在真實的數學與邏輯上明顯強過 8B 蒸餾。對多數推理工作來說,DeepSeek R1 Distill Llama 70B API 就是那個平衡點。

你透過 RouterBase、以 OpenAI chat-completions 協議呼叫 DeepSeek R1 Distill Llama 70B API;把模型設為 deepseek-r1-distill-llama-70b,客戶端就緒。DeepSeek R1 Distill Llama 70B API 按輸入輸出統一每百萬 token $0.80 計費,比標價低 5%,與覆蓋整個目錄的是同一把金鑰。

折中之道

DeepSeek R1 Distill Llama 70B API 為何是明智預設

足夠做真實推理的深度,價格又能一直開著。

均衡的蒸餾版

比 8B 更強,比 671B 更便宜更快——DeepSeek R1 Distill Llama 70B API 正是多數推理負載真正想要的折中之道。

R1 思維鏈

蒸餾自 DeepSeek-R1,DeepSeek R1 Distill Llama 70B API 逐步思考並串流輸出軌跡,讓你讀到推理,而不只是結果。

稠密 Llama 3.3 70B

稠密 70B 主幹帶來穩定延遲與強數學邏輯,且無需調度專家混合。

每百萬統一 $0.80

輸入輸出同價——DeepSeek R1 Distill Llama 70B API 按單一可預期費率計費,低於標價 5%,規模上易於預測。

開源權重

DeepSeek-R1-Distill-Llama-70B 為開源權重;你可自行評測,再讓 DeepSeek R1 Distill Llama 70B API 託管執行,無需自跑 GPU。

OpenAI 形態

一個 chat-completions 端點、一把 RouterBase 金鑰——指向 routerbase.com/v1、寫上模型名,跳過 DeepSeek SDK。

RouterBase dashboard preview
接上線

三步完成對 DeepSeek R1 Distill Llama 70B API 的首次呼叫

從金鑰到思維鏈,只需幾分鐘。

  1. 建立金鑰

    一把 RouterBase 金鑰即可存取 DeepSeek R1 Distill Llama 70B API 及目錄裡的每個模型。

  2. 指向並命名

    把任意 OpenAI 客戶端發往 routerbase.com/v1,model=deepseek/deepseek-r1-distill-llama-70b。串流與推理軌跡預設開啟。

  3. 讀取軌跡

    思維鏈與答案並排抵達;用量隨每個回應返回,成本一目了然。

定價

按使用量付費

RouterBase 透傳合作方價格,與模型官方公開 API 價格對比。

生產之選

把 DeepSeek R1 Distill Llama 70B API 設為預設的團隊

規模化執行的平衡點。

Camila RojasLead Engineer, Fathom

我們把 DeepSeek R1 Distill Llama 70B API 設為預設——它推理得像 R1,卻沒有 R1 的帳單。

Idris BelloCTO, Slate

8B 對我們的數學太輕,671B 又貴到不能常開。DeepSeek R1 Distill Llama 70B API 正好擊中中間。

Hana KimML Lead, Everline

稠密 70B 意味著可預期的延遲;DeepSeek R1 Distill Llama 70B API 從不給我們的 SLO 添驚嚇。

Viktor NovakStaff Engineer, Groundwork

輸入輸出統一 $0.80,預測變得毫不費力——月底不用再算進出比。

Renuka IyerFounder, Pathwise

DeepSeek R1 Distill Llama 70B API 串流輸出的思維鏈好到我們原樣保留。

Otto LindgrenPrincipal Engineer, Beacon Grid

開源權重讓我們拿它和 671B 對標,然後 DeepSeek R1 Distill Llama 70B API 讓我們上線更便宜的那個。

Zoe AlmeidaHead of AI, Rill

一把 RouterBase 金鑰,我們只改一個欄位就能把 DeepSeek R1 Distill Llama 70B API 和 Sonnet 做 A/B。

Jamal CarterBackend Lead, Overstory

我們把日常推理遷到 DeepSeek R1 Distill Llama 70B API,每任務成本下降,卻沒有一句品質抱怨。

Nina FalkEngineering Manager, Cindergrid

它現在是我們最先伸手去拿的推理模型——深度夠,價格合理。

DeepSeek R1 Distill Llama 70B API —— 常見問題

把它設為預設前該權衡的。

RouterBase 對 DeepSeek-R1-Distill-Llama-70B 的託管接入——把 DeepSeek-R1 推理蒸餾進稠密 Llama 3.3 70B 的模型,透過相容 OpenAI 的端點提供。