我們把 DeepSeek R1 Distill Llama 70B API 設為預設——它推理得像 R1,卻沒有 R1 的帳單。
Loading model information...
DeepSeek R1 Distill Llama 70B API —— R1 級推理,沒有 671B 的帳單
DeepSeek R1 Distill Llama 70B API 把 DeepSeek-R1 的思維鏈蒸餾進稠密的 Llama 3.3 70B——當 8B 太輕、完整 671B 又超出所需時的折中之道。
在小蒸餾與 6710 億參數旗艦之間,坐著那個明智的預設項。DeepSeek R1 Distill Llama 70B API 是把 DeepSeek-R1 推理蒸餾進稠密 Llama 3.3 70B 的模型:它保留讓 R1 值得用的逐步思維鏈,比完整模型更快、便宜得多,又在真實的數學與邏輯上明顯強過 8B 蒸餾。對多數推理工作來說,DeepSeek R1 Distill Llama 70B API 就是那個平衡點。
你透過 RouterBase、以 OpenAI chat-completions 協議呼叫 DeepSeek R1 Distill Llama 70B API;把模型設為 deepseek-r1-distill-llama-70b,客戶端就緒。DeepSeek R1 Distill Llama 70B API 按輸入輸出統一每百萬 token $0.80 計費,比標價低 5%,與覆蓋整個目錄的是同一把金鑰。
DeepSeek R1 Distill Llama 70B API 為何是明智預設
足夠做真實推理的深度,價格又能一直開著。
均衡的蒸餾版
比 8B 更強,比 671B 更便宜更快——DeepSeek R1 Distill Llama 70B API 正是多數推理負載真正想要的折中之道。
R1 思維鏈
蒸餾自 DeepSeek-R1,DeepSeek R1 Distill Llama 70B API 逐步思考並串流輸出軌跡,讓你讀到推理,而不只是結果。
稠密 Llama 3.3 70B
稠密 70B 主幹帶來穩定延遲與強數學邏輯,且無需調度專家混合。
每百萬統一 $0.80
輸入輸出同價——DeepSeek R1 Distill Llama 70B API 按單一可預期費率計費,低於標價 5%,規模上易於預測。
開源權重
DeepSeek-R1-Distill-Llama-70B 為開源權重;你可自行評測,再讓 DeepSeek R1 Distill Llama 70B API 託管執行,無需自跑 GPU。
OpenAI 形態
一個 chat-completions 端點、一把 RouterBase 金鑰——指向 routerbase.com/v1、寫上模型名,跳過 DeepSeek SDK。

三步完成對 DeepSeek R1 Distill Llama 70B API 的首次呼叫
從金鑰到思維鏈,只需幾分鐘。
建立金鑰
一把 RouterBase 金鑰即可存取 DeepSeek R1 Distill Llama 70B API 及目錄裡的每個模型。
指向並命名
把任意 OpenAI 客戶端發往 routerbase.com/v1,model=deepseek/deepseek-r1-distill-llama-70b。串流與推理軌跡預設開啟。
讀取軌跡
思維鏈與答案並排抵達;用量隨每個回應返回,成本一目了然。
按使用量付費
RouterBase 透傳合作方價格,與模型官方公開 API 價格對比。
把 DeepSeek R1 Distill Llama 70B API 設為預設的團隊
規模化執行的平衡點。
8B 對我們的數學太輕,671B 又貴到不能常開。DeepSeek R1 Distill Llama 70B API 正好擊中中間。
稠密 70B 意味著可預期的延遲;DeepSeek R1 Distill Llama 70B API 從不給我們的 SLO 添驚嚇。
輸入輸出統一 $0.80,預測變得毫不費力——月底不用再算進出比。
DeepSeek R1 Distill Llama 70B API 串流輸出的思維鏈好到我們原樣保留。
開源權重讓我們拿它和 671B 對標,然後 DeepSeek R1 Distill Llama 70B API 讓我們上線更便宜的那個。
一把 RouterBase 金鑰,我們只改一個欄位就能把 DeepSeek R1 Distill Llama 70B API 和 Sonnet 做 A/B。
我們把日常推理遷到 DeepSeek R1 Distill Llama 70B API,每任務成本下降,卻沒有一句品質抱怨。
它現在是我們最先伸手去拿的推理模型——深度夠,價格合理。
DeepSeek R1 Distill Llama 70B API —— 常見問題
把它設為預設前該權衡的。
RouterBase 對 DeepSeek-R1-Distill-Llama-70B 的託管接入——把 DeepSeek-R1 推理蒸餾進稠密 Llama 3.3 70B 的模型,透過相容 OpenAI 的端點提供。