DeepSeek V4.1 Flash
DeepSeek released a new Flash model with native visual understanding and improvements to agent workflows.
For buildersTest screenshots and document images alongside the text-only tasks in your evaluation set.
DeepSeek model announcements from December 2024 to September 2026, newest first. Compare what changed and follow the official sources.
Dates refer to the linked announcement or release milestone. Previews are labeled in the entries; historical availability and pricing may have changed.
DeepSeek released a new Flash model with native visual understanding and improvements to agent workflows.
For buildersTest screenshots and document images alongside the text-only tasks in your evaluation set.
DeepSeek released open-weight V4 preview models with one-million-token context and thinking/non-thinking modes.
For buildersCompare the Pro and Flash variants on your own long-input tasks.
DeepSeek released reasoning-first models for agent workflows, with a higher-reasoning Speciale variant on a temporary endpoint.
For buildersDifferent reasoning tiers should be compared on successful task completion and cost.
An experimental model introduced DeepSeek Sparse Attention for more efficient long-context processing.
For buildersLong-context efficiency matters for repeated analysis of large documents or codebases.
DeepSeek introduced a hybrid model combining thinking and non-thinking modes, with stronger agent capabilities.
For buildersUse reasoning selectively for the steps that require planning or difficult decisions.
An updated R1 checkpoint improved reasoning and added support for structured outputs and function calling.
For buildersEvaluate a reasoning workflow that must return valid structured data.
DeepSeek released its reasoning model and six smaller distilled models with open weights.
For buildersCompare reasoning quality across model sizes before choosing a deployment approach.
DeepSeek released a large mixture-of-experts language model with open weights.
For buildersA milestone for comparing hosted APIs with independently deployable language models.