DeepSeek V4.1 Flash — 快速、低價、支援圖片的推理模型
DeepSeek V4.1 Flash 是 DeepSeek 推出的快速、低價模型,支援 100 萬 token 上下文與圖片輸入。尖峰時段每百萬 token 價格:輸入 $0.24 / 1M、輸出 $0.96 / 1M;離峰時段半價。
- 思考可開可關
- 100 萬 token 上下文
- 圖片輸入
- 離峰時段價格
- 串流輸出
試用 DeepSeek V4.1 Flash
與 DeepSeek V4.1 Flash 對話、附加圖片,並查看每個回答的 token 數與費用。
deepseek-v4.1-flash$0.24 輸入 · $0.96 輸出,每百萬 token (目前為尖峰時段)
開啟思考時,temperature 不會生效,低於 0.95 的 top_p 會以 0.95 執行。
參數會被接受,但不會回傳對數機率。
DeepSeek V4.1 Flash 的真實輸出
一則提示詞,在關閉思考的情況下送給 DeepSeek V4.1 Flash。未經編輯。
Explain the difference between a mutex and a semaphore in two short paragraphs, and say when to use each.
A mutex is a locking mechanism designed for mutual exclusion: only the thread that locked it may unlock it, and it has a notion of ownership. It's binary (locked/unlocked) and is used to protect a shared resource so that only one thread accesses it at a time. A semaphore is a signaling mechanism built around a counter: it tracks how many permits are available, any thread can wait (decrement) or post (increment) it, and it doesn't care which thread does what. A binary semaphore looks similar to a mutex, but lacks ownership, so it can be released by a different thread—useful for signaling between threads rather than guarding a critical section. Use a mutex when you need to protect a shared resource (a data structure, file, or device) from concurrent access, which is the common case for locking critical sections. Use a semaphore when you need to control access to a pool of N identical resources (e.g., limiting concurrent database connections) or to signal between threads, such as a producer notifying a consumer that data is ready or a task-completion event. As a rule of thumb: mutex for mutual exclusion, semaphore for counting or signaling.
產生於 2026-09-28。每次執行的回答都可能不同。
DeepSeek V4.1 Flash 價格
DeepSeek V4.1 Flash 按 token 計費,從你的餘額扣款,並依請求所在時段適用尖峰或離峰費率。下方為即時費率,與你的帳戶實際扣款的費率相同。
範例用量僅供說明,以尖峰時段費率計算。實際帳單以每個請求回傳的 token 數及其執行時段為準。
| 項目 | 計為 | 費率 | 計費 |
|---|---|---|---|
| 輸入,尖峰時段 | 未命中快取的提示 token,UTC 週一至週五 01:00–04:00 與 06:00–10:00 | $0.24 / 1M | 每 100 萬 token |
| 輸出,尖峰時段 | 尖峰時段的回答 token 與推理 token | $0.96 / 1M | 每 100 萬 token |
| 快取輸入,尖峰時段 | 尖峰時段命中快取的提示 token | $0.0048 / 1M | 每 100 萬 token |
| 輸入,離峰時段 | 未命中快取的提示 token,其餘所有時段,含週末 | $0.12 / 1M | 每 100 萬 token |
| 輸出,離峰時段 | 離峰時段的回答 token 與推理 token | $0.48 / 1M | 每 100 萬 token |
| 快取輸入,離峰時段 | 離峰時段命中快取的提示 token | $0.0024 / 1M | 每 100 萬 token |
| 失敗的請求 | 任何回傳錯誤的請求 | 免費 | 不計費 |
DeepSeek V4.1 Flash 是什麼?
DeepSeek V4.1 Flash 是 DeepSeek 推出的快速、低價模型,於 2026 年 9 月 10 日發布。它是擁有 5,520 億參數的混合專家(MoE)模型,讀取時啟用 80 億參數、生成時啟用 160 億參數,原生支援圖片輸入,最多可接收 100 萬 token 的上下文。
DeepSeek V4.1 Flash 預設會先思考再回答,推理強度為 high。簡單快速的步驟可以關閉思考,難題則把推理強度調到 max。在 DeepSeek 自家的 API 中,這個模型稱為 deepseek-flash。
模型 ID:deepseek-v4.1-flash · 輸入:文字、圖片 · 輸出:文字 · 介面格式:Chat Completions、Responses、Anthropic Messages。
使用 DeepSeek V4.1 Flash 的兩種方式
從你的程式碼呼叫它,或交給你的程式開發代理代勞。
從後端呼叫
把 OpenAI SDK 指向 SeedRouter,程式碼不用改。DeepSeek V4.1 Flash 支援 Chat Completions、Responses 與 Anthropic Messages 三種格式。
- 1建立 API 金鑰
- 2把 base URL 設為 https://api.seedrouter.ai/v1
- 3將 model 設為 deepseek-v4.1-flash
交給你的程式開發代理
複製現成的提示。你的代理會在產生任何費用前,先顯示請求內容和費用。
- 1複製提示
- 2貼到你的代理中
- 3核准請求
DeepSeek V4.1 Flash 擅長什麼
DeepSeek V4.1 Flash 專為重視速度與費用的大量工作打造。
代理迴圈
DeepSeek V4.1 Flash 能以低廉的價格執行大量步驟,快取的上下文讓重複讀取保持低價。
日常程式開發
在 Codex、Claude Code 或你自己的工具中,用 DeepSeek V4.1 Flash 撰寫、審查與修正程式碼。
提示中附上圖片
DeepSeek V4.1 Flash 能讀懂透過 URL 或 base64 傳入的截圖、圖表與照片。
結構化輸出
請 DeepSeek V4.1 Flash 輸出 json_object,直接解析回答。
思考可開可關
簡單問題可略過 DeepSeek V4.1 Flash 的推理快速回答,難題則調高推理強度。
離峰時段費率
在尖峰時段以外執行的批次作業,只需支付尖峰價格的一半。
團隊用 DeepSeek V4.1 Flash 打造什麼
大規模分類
關閉思考,用 DeepSeek V4.1 Flash 為工單、評論與日誌加上標籤,費用最低。
程式開發代理
在 Codex 中執行 DeepSeek V4.1 Flash,打造快速、低價的「編輯—執行」迴圈。
截圖辨識
把截圖與圖表轉成文字或 JSON。
四個步驟開始使用 DeepSeek V4.1 Flash
- 01
建立金鑰
登入並建立 API 金鑰。需要時隨時新增額度;沒有訂閱。
- 02
改變 base URL
把 OpenAI SDK 指向 https://api.seedrouter.ai/v1,程式碼維持不變。
- 03
選擇 DeepSeek V4.1 Flash
將 model 設為 deepseek-v4.1-flash。可依請求關閉思考或設定推理強度。
- 04
檢查帳單
每個請求在你的用量記錄中顯示其 token 和費用。
SeedRouter 上的 DeepSeek V4.1 Flash 一覽
你今天可以使用的內容。
| 功能 | DeepSeek V4.1 Flash |
|---|---|
| Chat Completions | 支援,官方格式 |
| Responses | 是,可搭配 Codex 使用 |
| Anthropic Messages | 是,可搭配 Claude Code 使用 |
| 圖片輸入 | 是,經由 URL 或 base64 |
| 串流輸出 | 是 |
| 上下文快取 | 是,自動生效,命中快取以較低費率計費 |
| JSON 輸出 | 是,json_object |
| 失敗的請求 | 不計費 |
不提供 FIM 與對話前綴續寫(beta)、Files API 與對數機率。
DeepSeek V4.1 Flash 的使用限制
100 萬 token 上下文
在 DeepSeek V4.1 Flash 上,提示、圖片與對話紀錄共用一個 100 萬 token 的視窗。
384K token 輸出
單次 DeepSeek V4.1 Flash 請求最多輸出 393,216 token,包含推理內容。
開啟思考時的取樣參數
開啟思考時,temperature 不會生效,top_p 維持在 0.95 或以上。
圖片大小
DeepSeek V4.1 Flash 的圖片 URL 所指向的檔案最大可達 32 MiB。
為什麼透過 SeedRouter 呼叫 DeepSeek V4.1 Flash
一把金鑰,三種格式
用同一把金鑰,就能以 Chat Completions、Responses 或 Anthropic Messages 呼叫 DeepSeek V4.1 Flash。
按 token 付費
儲值一次,所有模型都能用。沒有方案,沒有月費。
失敗不計費
回傳錯誤的請求不計費。
即時價格
這頁上的尖峰與離峰時段費率,就是你的帳戶付的費率。
上下文快取
重複的提示前綴會從快取讀取,以較低的 DeepSeek V4.1 Flash 費率計費。
回覆中附帶推理內容
開啟思考時,推理內容會在回答旁的 reasoning_content 中回傳。
串流輸出
設定 stream: true,就能邊產生邊接收 DeepSeek V4.1 Flash 的 token。
相關模型
你可以用同一把金鑰呼叫的其他模型。
用 OpenAI 或 Anthropic SDK 串接 Kimi K3:100 萬 token 上下文,推理始終開啟,強度可選 low、high 或 max,按 token 即時計價,還有 Playground 可線上試用。
檢視價格DeepSeek V4.1 Flash 使用指南
此模型的逐步說明和比較。
DeepSeek V4.1 Flash API 串接教學:取得 API 金鑰,用 OpenAI SDK 呼叫,開啟或關閉思考模式、串流輸出、傳送圖片,並修正新手常遇到的錯誤。
DeepSeek V4.1 Flash API 每百萬 token 價格:尖峰時段與離峰時段費率、快取命中、各費率適用的時間、計費公式,以及一次請求的實際費用。
如何透過 SeedRouter 把 DeepSeek V4.1 Flash 設為 Claude Code 與 OpenAI Codex CLI 的模型:環境變數、config.toml 設定,以及我們的實測結果。
DeepSeek V4.1 Flash 與 DeepSeek V4 Pro 比較:每 token 價格、圖片輸入、並行上限與 DeepSeek 公布的基準測試成績,並給出各種工作的明確選擇。
從 DeepSeek V4 Flash 0731 到 V4.1 Flash 改了什麼:全新架構、原生視覺、更小的快取、更低的價格、基準測試,以及舊模型名稱的去向。
DeepSeek V4.1 Flash 免費嗎?權重可免費下載,API 以 token 低價計費,這裡告訴你今天如何幾乎零成本地試用它。
DeepSeek V4.1 Flash 常見問題
DeepSeek V4.1 Flash 是什麼?+
DeepSeek V4.1 Flash 是 DeepSeek 推出的快速、低價模型,於 2026 年 9 月 10 日發布。它是擁有 5,520 億參數的混合專家(MoE)模型,原生支援圖片輸入,具備 100 萬 token 的上下文視窗,思考可開可關。
DeepSeek V4.1 Flash 多少錢?+
DeepSeek V4.1 Flash 在尖峰時段(UTC 週一至週五 01:00–04:00 與 06:00–10:00)按每百萬 token 輸入 $0.24 / 1M、輸出 $0.96 / 1M 計費。其餘所有時段皆為離峰時段,費率減半;命中快取的價格只是輸入的一小部分。
DeepSeek V4.1 Flash 可以免費使用嗎?+
每個新註冊的 SeedRouter 帳戶都附有 $0.10 的免費餘額,足以送出許多簡短請求來試用 DeepSeek V4.1 Flash。之後按 token 付費,沒有訂閱,失敗的請求不收費。
DeepSeek V4.1 Flash 支援圖片嗎?+
支援。DeepSeek V4.1 Flash 原生就能讀取圖片,可在請求中以公開 URL 或 base64 傳入;它回傳文字。
如何關閉 DeepSeek V4.1 Flash 的思考?+
把 thinking 設為 disabled,或把 reasoning_effort 設為 none。這樣回答會立即回傳,並消耗較少的輸出 token。
DeepSeek V4.1 Flash 和 DeepSeek V4 Flash 是同一個模型嗎?+
不是。DeepSeek V4.1 Flash 於 2026 年 9 月 10 日取代了 V4 Flash,採用新架構並原生支援圖片輸入。DeepSeek 已停用 V4 Flash,並把它的舊模型名稱導向 DeepSeek V4.1 Flash。
DeepSeek V4.1 Flash 可以在本機執行嗎?+
它的權重已發布在 Hugging Face 上。它擁有 5,520 億參數,需要多 GPU 硬體才能執行,因此大多數團隊透過 API 呼叫 DeepSeek V4.1 Flash。
DeepSeek V4.1 Flash 的模型 ID 是什麼?要怎麼串接?+
在 SeedRouter 上,模型 ID 是 deepseek-v4.1-flash。建立 SeedRouter API 金鑰,把 OpenAI SDK 指向 https://api.seedrouter.ai/v1,再將 model 設為 deepseek-v4.1-flash 送出請求。
DeepSeek V4.1 Flash 請求失敗會收費嗎?+
不會。回傳錯誤的 DeepSeek V4.1 Flash 請求不收費。你只需為成功完成的請求所回報的 token 付費。
立即試用 DeepSeek V4.1 Flash
幾分鐘內,就能從 Playground 或你自己的程式碼送出第一個 DeepSeek V4.1 Flash 請求。
