Text generation / DeepSeek
DeepSeek-V4-Flash-0731
API name: deepseek-v4-flash-0731
Deep Thinking
Overview
A flagship MoE large model with 1.6 trillion parameters and 49 billion activated parameters, natively supporting context lengths of up to one million tokens. Trained on a vast corpus of high-quality data, it excels in advanced mathematical reasoning, complex logical inference, specialized coding, and deep analysis of long-form text, making it well-suited for demanding applications such as cutting-edge research, sophisticated office workflows, and advanced AI agents.
- Model type: Text generation
- Input: Text
- Output: Text
- Endpoints: chat / chat_completion
Tiered billing details
| Tier | Input | Output | Cache hit |
|---|---|---|---|
| Off-peak | $0.0002/1K | $0.0006/1K | $0.0000/1K |
| Peak | $0.0004/1K | $0.0013/1K | $0.0000/1K |
Billed bylen (input context token count). Coefficient is the $ / 1M tokens price. Default group price (other groups converted by discount).
Code example
Call with deepseek-v4-flash-0731 :
curl https://www.starunion.net/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek-v4-flash-0731",
"messages": [{ "role": "user", "content": "Hello" }]
}'