Skip to main content
DeepSeek V4.1 Flash 通过 AnyFast 提供服务,公开模型 ID 为 deepseek-v4.1-flash,调用端点为 OpenAI 兼容的 POST /v1/chat/completions。支持文本和图片输入、思考与非思考模式、函数调用、JSON 输出、流式输出和提示缓存。
DeepSeek 官方将该版本命名为 DeepSeek-V4.1-Flash,并在官方 API 中使用 deepseek-flash。调用 AnyFast 时请使用 deepseek-v4.1-flash。

模型规格

快速示例

cURL

思考模式

思考模式默认开启,默认强度为 high。对于不需要推理、更加关注延迟的请求,可以使用 {"thinking":{"type":"disabled"}}。开启思考时,推理内容位于 choices[0].message.reasoning_content,最终答案位于 choices[0].message.content。 官方模型支持 low、high、max 三档思考强度。思考模式下,temperature、presence_penalty 和 frequency_penalty 不生效。请求包含 tools 时,后续工具调用轮次需要完整回传先前 assistant 消息中的 reasoning_content。

图片输入

image_url 内容块只能放在 user 消息中。上游支持 JPEG、PNG、GIF 和 WebP。URL 必须可以直接通过 HTTP(S) 下载;本地图片可以使用 Base64 data URL。图片 Token 计入 usage.prompt_tokens。

常用参数

DeepSeek V4.1 Flash API 参考

查看完整的 Chat Completions 请求与响应字段。
资料来源:DeepSeek V4.1 Flash 发布说明、模型与价格、思考模式和图片理解。