Large capacity, sparse activation
V4 Flash has 284B total parameters while activating 13B per token. DeepSeek positions it as the speed-and-economy member of V4, with reasoning close to Pro and comparable results on simple agent tasks.
- Current published version: DeepSeek-V4-Flash-0731
- Thinking and non-thinking modes in one model
- JSON output, tool calls, and long-context workflows