Claude Opus 5.5 已在 SeedRouter 上線
DeepSeek文字生成

DeepSeek V4.1 Flash — 快速、低價、支援圖片的推理模型

DeepSeek V4.1 Flash 是 DeepSeek 推出的快速、低價模型,支援 100 萬 token 上下文與圖片輸入。尖峰時段每百萬 token 價格:輸入 $0.24 / 1M、輸出 $0.96 / 1M;離峰時段半價。

可用$0.24 / 1M 輸入 · $0.96 / 1M 輸出$0.30 / 1M 輸入 · $1.20 / 1M 輸出-20%
  • 思考可開可關
  • 100 萬 token 上下文
  • 圖片輸入
  • 離峰時段價格
  • 串流輸出
上下文視窗
100 萬 token
最大輸出
384K token
推理強度等級
none、low、high、max
發布日期
2026 年 9 月 10 日

試用 DeepSeek V4.1 Flash

與 DeepSeek V4.1 Flash 對話、附加圖片,並查看每個回答的 token 數與費用。

參數

deepseek-v4.1-flash$0.24 輸入 · $0.96 輸出,每百萬 token (目前為尖峰時段)

思考
Reasoning effort

開啟思考時,temperature 不會生效,低於 0.95 的 top_p 會以 0.95 執行。

Log probabilities

參數會被接受,但不會回傳對數機率。

對話總計$0.000000
就緒

試試下面這幾條提示詞

訊息按下列費率從你的餘額計費,就像 API 呼叫一樣。失敗的請求不收費。

DeepSeek V4.1 Flash 的真實輸出

一則提示詞,在關閉思考的情況下送給 DeepSeek V4.1 Flash。未經編輯。

Explain the difference between a mutex and a semaphore in two short paragraphs, and say when to use each.

A mutex is a locking mechanism designed for mutual exclusion: only the thread that locked it may unlock it, and it has a notion of ownership. It's binary (locked/unlocked) and is used to protect a shared resource so that only one thread accesses it at a time. A semaphore is a signaling mechanism built around a counter: it tracks how many permits are available, any thread can wait (decrement) or post (increment) it, and it doesn't care which thread does what. A binary semaphore looks similar to a mutex, but lacks ownership, so it can be released by a different thread—useful for signaling between threads rather than guarding a critical section. Use a mutex when you need to protect a shared resource (a data structure, file, or device) from concurrent access, which is the common case for locking critical sections. Use a semaphore when you need to control access to a pool of N identical resources (e.g., limiting concurrent database connections) or to signal between threads, such as a producer notifying a consumer that data is ready or a task-completion event. As a rule of thumb: mutex for mutual exclusion, semaphore for counting or signaling.

產生於 2026-09-28。每次執行的回答都可能不同。

DeepSeek V4.1 Flash 價格

DeepSeek V4.1 Flash 按 token 計費,從你的餘額扣款,並依請求所在時段適用尖峰或離峰費率。下方為即時費率,與你的帳戶實際扣款的費率相同。

一個典型的請求2,000 輸入 token,1,000 輸出 token,尖峰時段
$0.0014每個請求
在離峰時段,同樣的請求只要一半費用。
$20.00 能買多少以尖峰時段費率輸出 token
20,833,333輸出 token
離峰時段可買到兩倍的量。快取輸入的價格只是輸入的一小部分。

範例用量僅供說明,以尖峰時段費率計算。實際帳單以每個請求回傳的 token 數及其執行時段為準。

項目計為費率計費
輸入,尖峰時段未命中快取的提示 token,UTC 週一至週五 01:00–04:00 與 06:00–10:00$0.24 / 1M每 100 萬 token
輸出,尖峰時段尖峰時段的回答 token 與推理 token$0.96 / 1M每 100 萬 token
快取輸入,尖峰時段尖峰時段命中快取的提示 token$0.0048 / 1M每 100 萬 token
輸入,離峰時段未命中快取的提示 token,其餘所有時段,含週末$0.12 / 1M每 100 萬 token
輸出,離峰時段離峰時段的回答 token 與推理 token$0.48 / 1M每 100 萬 token
快取輸入,離峰時段離峰時段命中快取的提示 token$0.0024 / 1M每 100 萬 token
失敗的請求任何回傳錯誤的請求免費不計費

DeepSeek V4.1 Flash 是什麼?

DeepSeek V4.1 Flash 是 DeepSeek 推出的快速、低價模型,於 2026 年 9 月 10 日發布。它是擁有 5,520 億參數的混合專家(MoE)模型,讀取時啟用 80 億參數、生成時啟用 160 億參數,原生支援圖片輸入,最多可接收 100 萬 token 的上下文。

DeepSeek V4.1 Flash 預設會先思考再回答,推理強度為 high。簡單快速的步驟可以關閉思考,難題則把推理強度調到 max。在 DeepSeek 自家的 API 中,這個模型稱為 deepseek-flash。

模型 ID:deepseek-v4.1-flash · 輸入:文字、圖片 · 輸出:文字 · 介面格式:Chat Completions、Responses、Anthropic Messages。

模型 ID
deepseek-v4.1-flash
上下文視窗
100 萬 token
最大輸出
384K token
推理強度等級
none、low、high、max
輸入
文字、圖片

使用 DeepSeek V4.1 Flash 的兩種方式

從你的程式碼呼叫它,或交給你的程式開發代理代勞。

API

從後端呼叫

最適合應用程式和服務

把 OpenAI SDK 指向 SeedRouter,程式碼不用改。DeepSeek V4.1 Flash 支援 Chat Completions、Responses 與 Anthropic Messages 三種格式。

  1. 1建立 API 金鑰
  2. 2把 base URL 設為 https://api.seedrouter.ai/v1
  3. 3將 model 設為 deepseek-v4.1-flash
代理

交給你的程式開發代理

最適合 Codex、Claude Code 和其他代理

複製現成的提示。你的代理會在產生任何費用前,先顯示請求內容和費用。

  1. 1複製提示
  2. 2貼到你的代理中
  3. 3核准請求

DeepSeek V4.1 Flash 擅長什麼

DeepSeek V4.1 Flash 專為重視速度與費用的大量工作打造。

代理迴圈

DeepSeek V4.1 Flash 能以低廉的價格執行大量步驟,快取的上下文讓重複讀取保持低價。

日常程式開發

在 Codex、Claude Code 或你自己的工具中,用 DeepSeek V4.1 Flash 撰寫、審查與修正程式碼。

提示中附上圖片

DeepSeek V4.1 Flash 能讀懂透過 URL 或 base64 傳入的截圖、圖表與照片。

結構化輸出

請 DeepSeek V4.1 Flash 輸出 json_object,直接解析回答。

思考可開可關

簡單問題可略過 DeepSeek V4.1 Flash 的推理快速回答,難題則調高推理強度。

離峰時段費率

在尖峰時段以外執行的批次作業,只需支付尖峰價格的一半。

團隊用 DeepSeek V4.1 Flash 打造什麼

大規模分類

關閉思考,用 DeepSeek V4.1 Flash 為工單、評論與日誌加上標籤,費用最低。

程式開發代理

在 Codex 中執行 DeepSeek V4.1 Flash,打造快速、低價的「編輯—執行」迴圈。

截圖辨識

把截圖與圖表轉成文字或 JSON。

四個步驟開始使用 DeepSeek V4.1 Flash

  1. 01

    建立金鑰

    登入並建立 API 金鑰。需要時隨時新增額度;沒有訂閱。

  2. 02

    改變 base URL

    把 OpenAI SDK 指向 https://api.seedrouter.ai/v1,程式碼維持不變。

  3. 03

    選擇 DeepSeek V4.1 Flash

    將 model 設為 deepseek-v4.1-flash。可依請求關閉思考或設定推理強度。

  4. 04

    檢查帳單

    每個請求在你的用量記錄中顯示其 token 和費用。

取得 API 金鑰

SeedRouter 上的 DeepSeek V4.1 Flash 一覽

你今天可以使用的內容。

功能DeepSeek V4.1 Flash
Chat Completions支援,官方格式
Responses是,可搭配 Codex 使用
Anthropic Messages是,可搭配 Claude Code 使用
圖片輸入是,經由 URL 或 base64
串流輸出是
上下文快取是,自動生效,命中快取以較低費率計費
JSON 輸出是,json_object
失敗的請求不計費

不提供 FIM 與對話前綴續寫(beta)、Files API 與對數機率。

DeepSeek V4.1 Flash 的使用限制

100 萬 token 上下文

在 DeepSeek V4.1 Flash 上,提示、圖片與對話紀錄共用一個 100 萬 token 的視窗。

384K token 輸出

單次 DeepSeek V4.1 Flash 請求最多輸出 393,216 token,包含推理內容。

開啟思考時的取樣參數

開啟思考時,temperature 不會生效,top_p 維持在 0.95 或以上。

圖片大小

DeepSeek V4.1 Flash 的圖片 URL 所指向的檔案最大可達 32 MiB。

為什麼透過 SeedRouter 呼叫 DeepSeek V4.1 Flash

一把金鑰,三種格式

用同一把金鑰,就能以 Chat Completions、Responses 或 Anthropic Messages 呼叫 DeepSeek V4.1 Flash。

按 token 付費

儲值一次,所有模型都能用。沒有方案,沒有月費。

失敗不計費

回傳錯誤的請求不計費。

即時價格

這頁上的尖峰與離峰時段費率,就是你的帳戶付的費率。

上下文快取

重複的提示前綴會從快取讀取,以較低的 DeepSeek V4.1 Flash 費率計費。

回覆中附帶推理內容

開啟思考時,推理內容會在回答旁的 reasoning_content 中回傳。

串流輸出

設定 stream: true,就能邊產生邊接收 DeepSeek V4.1 Flash 的 token。

相關模型

你可以用同一把金鑰呼叫的其他模型。

Kimi K3
kimi-k3

用 OpenAI 或 Anthropic SDK 串接 Kimi K3:100 萬 token 上下文,推理始終開啟,強度可選 low、high 或 max,按 token 即時計價,還有 Playground 可線上試用。

檢視價格
claude-fable-5
claude-fable-5

Anthropic

檢視價格
claude-fable-5-1
claude-fable-5-1

Anthropic

檢視價格
claude-opus-5-5
claude-opus-5-5

Anthropic

檢視價格
dreamina-seedance-2-0
dreamina-seedance-2-0

ByteDance

檢視價格
dreamina-seedance-2-0-fast
dreamina-seedance-2-0-fast

ByteDance

檢視價格

DeepSeek V4.1 Flash 使用指南

此模型的逐步說明和比較。

DeepSeek V4.1 Flash 常見問題

DeepSeek V4.1 Flash 是什麼?+

DeepSeek V4.1 Flash 是 DeepSeek 推出的快速、低價模型,於 2026 年 9 月 10 日發布。它是擁有 5,520 億參數的混合專家(MoE)模型,原生支援圖片輸入,具備 100 萬 token 的上下文視窗,思考可開可關。

DeepSeek V4.1 Flash 多少錢?+

DeepSeek V4.1 Flash 在尖峰時段(UTC 週一至週五 01:00–04:00 與 06:00–10:00)按每百萬 token 輸入 $0.24 / 1M、輸出 $0.96 / 1M 計費。其餘所有時段皆為離峰時段,費率減半;命中快取的價格只是輸入的一小部分。

DeepSeek V4.1 Flash 可以免費使用嗎?+

每個新註冊的 SeedRouter 帳戶都附有 $0.10 的免費餘額,足以送出許多簡短請求來試用 DeepSeek V4.1 Flash。之後按 token 付費,沒有訂閱,失敗的請求不收費。

DeepSeek V4.1 Flash 支援圖片嗎?+

支援。DeepSeek V4.1 Flash 原生就能讀取圖片,可在請求中以公開 URL 或 base64 傳入;它回傳文字。

如何關閉 DeepSeek V4.1 Flash 的思考?+

把 thinking 設為 disabled,或把 reasoning_effort 設為 none。這樣回答會立即回傳,並消耗較少的輸出 token。

DeepSeek V4.1 Flash 和 DeepSeek V4 Flash 是同一個模型嗎?+

不是。DeepSeek V4.1 Flash 於 2026 年 9 月 10 日取代了 V4 Flash,採用新架構並原生支援圖片輸入。DeepSeek 已停用 V4 Flash,並把它的舊模型名稱導向 DeepSeek V4.1 Flash。

DeepSeek V4.1 Flash 可以在本機執行嗎?+

它的權重已發布在 Hugging Face 上。它擁有 5,520 億參數,需要多 GPU 硬體才能執行,因此大多數團隊透過 API 呼叫 DeepSeek V4.1 Flash。

DeepSeek V4.1 Flash 的模型 ID 是什麼?要怎麼串接?+

在 SeedRouter 上,模型 ID 是 deepseek-v4.1-flash。建立 SeedRouter API 金鑰,把 OpenAI SDK 指向 https://api.seedrouter.ai/v1,再將 model 設為 deepseek-v4.1-flash 送出請求。

DeepSeek V4.1 Flash 請求失敗會收費嗎?+

不會。回傳錯誤的 DeepSeek V4.1 Flash 請求不收費。你只需為成功完成的請求所回報的 token 付費。

立即試用 DeepSeek V4.1 Flash

幾分鐘內,就能從 Playground 或你自己的程式碼送出第一個 DeepSeek V4.1 Flash 請求。