DeepSeek released V4-Flash 0731, a new model variant, on July 31, 2026. The Latent Space AINews digest notes this as the sole significant development of the day. The release extends DeepSeek's V4 model family with a Flash-tier (likely faster/cheaper) variant.
DeepSeek has released an update to DeepSeek-V4-Flash, a faster/lighter variant of their flagship V4 model family. The announcement appeared in DeepSeek's official API documentation changelog. Community engagement on Hacker News (265 points, 112 comments) suggests meaningful practitioner interest in the update.
Simon Willison flagged the release of DeepSeek-V4-Flash-0731, a new checkpoint in the DeepSeek V4 Flash model line. The post is a brief link-log entry with minimal commentary. DeepSeek's Flash variants are typically smaller, faster inference-optimized models in their flagship family.
DeepSeek has published DeepSeek-V4-Flash-0731, a new text-generation model on Hugging Face under the deepseek_v4 model family. The release is tagged as endpoints-compatible with fp8 and 8-bit quantization support, suggesting an efficiency-oriented variant of the V4 series. With 640 likes at release and zero downloads logged, this appears to be a freshly published checkpoint in the V4 Flash line.
Artificial Analysis has published a benchmarking and pricing analysis of DeepSeek V4 Flash 0731, a new model variant from DeepSeek. The post is gaining significant traction on Hacker News with 343 points and 172 comments, suggesting notable community interest. The analysis covers intelligence benchmarks, performance metrics, and cost positioning for this flash-tier model.
DeepSeek has published a new model checkpoint, DeepSeek-V4-Flash-DSpark, on Hugging Face under the deepseek_v4 model family. The release is tagged as a text-generation model with FP8 and 8-bit support, suggesting an efficiency-optimized variant. The 'Flash' and 'DSpark' naming implies a faster or distilled derivative of the DeepSeek V4 flagship. Download counts are near zero, indicating a very recent upload.
DeepSeek has released DeepSeek-V4-Flash-Base, a new open-weights base model, on Hugging Face. The model uses FP8 precision and the deepseek_v4 architecture with safetensors format. Early traction is notable with over 66,000 downloads and 241 likes shortly after release, suggesting significant community interest in a 'Flash' variant of the V4 series.
DeepSeek has released DeepSeek-V4-Flash, a new text-generation model published on Hugging Face under the deepseek-ai organization. The model supports FP8 and 8-bit quantization and is tagged as conversational and endpoints-compatible. With over 2.8 million downloads and 1,455 likes, it has seen substantial early uptake.
DeepSeek has released DeepSeek-V4 as an open-weights preview, comprising two MoE variants: V4-Pro (1.6T total / 49B active parameters) and V4-Flash (284B total / 13B active parameters). Both models support 1M token context by default, enabled by a novel Token-wise compression and DeepSeek Sparse Attention (DSA) architecture. V4-Pro claims open-source SOTA on agentic coding benchmarks and world-class math/STEM/coding performance rivaling top closed-source models, while V4-Flash offers near-parity reasoning at lower cost and latency. The API is live today with OpenAI and Anthropic compatibility, and legacy model endpoints will be retired in July 2026.