Your privacy choices

Allow optional cookies for referral attribution, visit analytics, and Google Ads purchase measurement.

Model selection and API access

DeepSeek V4.1 Flash API pricing and integration guide

Check DeepSeek V4.1 Flash gateway pricing, token rates, documented 1M context and multimodal integration considerations.

Current prices and model reference

Model ID
deepseek-v4.1-flash
Documented context
1M
Documented image input
Supported
Live model multiplier
37.5×
Input price
¥15 / 1M tokens
Output price
¥15 / 1M tokens

Prices are in CNY. A 1× multiplier equals CNY 0.4 per million input tokens. Output and cache charges follow the billing configuration; your usage record is authoritative.

View the full live pricing catalog

When to consider this model

Consider this route for workflows combining Chinese documents, code and screenshots. The gateway reference lists 1M context and multimodal input; compare output quality, latency and recorded cost on your own tasks.

How to evaluate it

Validate a short text request first, then add one screenshot with a specific question. Increase document length gradually and keep a text-only control to check whether visual details were understood.

What to check before integration

Use the exact v4.1 model ID. deepseek-v4-flash is a separate route with its own capabilities and pricing. Image input also depends on the format your client sends.

Client configuration

In a client supporting a custom API endpoint, enter this base URL and model ID with your own API key. Select the protocol described in the client tutorial and validate a short request first.

Base URL: https://api.llm-token.cn/v1
Model ID: deepseek-v4.1-flash