Compare DeepSeek V4.1 Flash and DeepSeek V4 Pro 0813 on key metrics including price, context length, throughput, and other model features.
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on output from a 552B-parameter backbone, an asymmetric split that keeps per-token compute low relative to the model's total size. Image understanding is native to the architecture, with visual and text embeddings trained jointly from the start of pre-training rather than added afterward as in the earlier experimental V4 Flash Vision Exp.
DeepSeek V4 Pro 0813 is the official release of DeepSeek V4 Pro, superseding the preview version, with greatly enhanced agentic capabilities and performance improvements that are especially pronounced in production environments. It is built on the DeepSeek V4 Pro (Preview) model structure, with a DSpark speculative decoding module attached. The model natively supports a 1M-token context window and flexible reasoning effort: low for simple tasks, high for daily agent workflows, and max for complex ones.