XiaoZhan AIhubXiaoZhan AIhub
HomeAI ChatAgentsCreateCreative AgentModel catalogueTopicsPriceDocs
XiaoZhan AIhubXiaoZhan AIhub
HomeAI ChatAgentsCreateCreative AgentModel catalogueTopicsPriceDocs
XiaoZhan AIhubXiaoZhan AIhub
HomeAI ChatAgentsCreateCreative AgentModel catalogueTopicsPriceDocs
Model catalogueDeepSeek V4.1 Flash

DeepSeek V4.1 Flash

NEW

Model ID

Put this value in the request body's `model` field — not the display name above.

Text generationReasoningImage understandingToolsJSONPrefix1MOpen source

by DeepSeek · DeepSeek V4.1 Flash — 552B-parameter MoE (asymmetric Causal-Encoder-Decoder architecture), native multimodal vision understanding, 1M context / 384K max output, thinking mode by default, direct from the vendor.

Try it now

Quick start

EndpointPOST /v1/chat/completions
AuthenticationAuthorization: Bearer sk-gpushare-…Get API key
Call styleSynchronous: one request returns the result
Base URLhttps://usapi.dflop.top/v1
cURL
curl https://usapi.dflop.top/v1/chat/completions \
  -H "Authorization: Bearer sk-gpushare-xxxxxxxx" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4.1-flash",
    "messages": [
      {"role": "user", "content": "你可以帮我做什么"}
    ]
  }'

Pricing

Billed per token: input and output are priced separately and settled on the upstream's reported usage.

Rates
Input 0.12Credits/K·Output 0.48Credits/K
  • Input tokens that hit the prompt cache are billed at 2.4 Credits / 1M — cheaper than full-price input.
  • Failed requests (upstream 4xx / 5xx) are not billed.
  • All your API keys share one balance; when it runs out the API returns HTTP 402 `quota_exceeded`.

Sign in and any personal discount on your account will show up here. Sign in

Capabilities & specs

Category
Text generation
Context window
1M tokens
Default max output
128k tokens
Endpoint family
Chat Completions
Protocols
OpenAI Chat
Released
2026-09-10
License
Open source

Try it

Pick a key under "API access" and fire one real request to see the response — the estimated cost is shown before you send.

Try it in chat

Common errors

404 model_not_found

You passed the display name instead of the model ID.✗ DeepSeek V4.1 Flash→✓ deepseek-v4.1-flash

→ Copy the Model ID at the top of this page

402 quota_exceeded

Insufficient account balance. All your API keys share one balance.

→ Add funds in your account

401 / 403

The API key is invalid, or it has a model allowlist that doesn't include this ID.

→ Check the key's status and allowed models

400 invalid_request

A request body field is invalid, or this model doesn't support the protocol endpoint you called.

→ Check your request body against the sample code above

More models from this vendor

DeepSeek V4 Prodeepseek-v4-pro
Try it

关于 DeepSeek V4.1 Flash 的常见问题

怎么调用 DeepSeek V4.1 Flash?
请求体的 model 字段填 deepseek-v4.1-flash,打到 https://usapi.dflop.top/v1/chat/completions,把 sk-gpushare- 开头的密钥放进 Authorization: Bearer 请求头。OpenAI、Anthropic、Google Gemini 三套官方 SDK 都能直接调用,只需把 base_url 指向本网关。
DeepSeek V4.1 Flash 的 API 价格是多少?
  • ·输入:每 100 万 token 120 积分(≈ ¥2 / ≈ $0.297)
  • ·输出:每 100 万 token 480 积分(≈ ¥8 / ≈ $1.19)
  • ·缓存命中的输入:每 100 万 token 2.4 积分(≈ ¥0.04 / ≈ $0.0059)
  • ·平台以积分计价,60 积分 = ¥1;每次调用先按上限预扣,返回后按实际用量结算,差额自动退回。
DeepSeek V4.1 Flash 的上下文窗口有多大?
1,000,000 token,单次请求默认输出上限 128,000 token。
DeepSeek V4.1 Flash 支持哪些能力?
图片输入、工具调用、思考模式、JSON 模式、前缀续写、开源权重。兼容协议:OpenAI Chat Completions。

输入价格约合每 100 万 token $0.297。本页的 Markdown 版本:/models/deepseek-v4.1-flash.md