DeepSeek is a Chinese AI company founded in Hangzhou in 2023. Its first public model family, DeepSeek Coder, arrived on November 2, 2023. The company then expanded into general language models, mathematical reasoning, formal proof, vision, document recognition, and dedicated reasoning models.
DeepSeek-V4-Flash-Vision-Exp is the company’s newest model as of August 21, 2026. It is an experimental multimodal API model called through deepseek-v4-flash-vision-exp. It matches V4-Flash on text capabilities and accepts image input alongside text through the Chat Completions, Messages, and Responses APIs.
Between DeepSeek Coder and the V4-Flash-Vision-Exp update, the company repeatedly folded ideas from its specialist research into its main model line. It also expanded its developer platform with DeepSeek Harness, an open-source system for building and running AI agents. The dates below place these model, product, and API milestones in order.
Last Updated: August 21, 2026
Key DeepSeek Release Dates
| Milestone | Date and details |
|---|---|
| DeepSeek founded | 2023 Liang Wenfeng founded the company with financial backing from High-Flyer. |
| First public model | November 2, 2023 DeepSeek Coder was the company’s first public model family. |
| DeepSeek-V3 | December 26, 2024 V3 became the foundation for R1 and powered the DeepSeek App at launch. |
| DeepSeek App | January 15, 2025 The official App launched with DeepSeek-V3. |
| DeepSeek-R1 | January 20, 2025 R1 arrived five days after the App launched. |
| DeepSeek-V4 Preview | April 24, 2026 DeepSeek introduced the V4-Pro and V4-Flash previews. |
| DeepSeek-V4-Flash API public beta | July 31, 2026 A post-trained Flash update added more powerful agent performance, Responses API support, and Codex integration. |
| DeepSeek-V4-Pro-0813 API update | August 13, 2026 The current Pro version added native Responses API support alongside the existing Anthropic-compatible interface. |
| DeepSeek Harness v0.1 | August 13, 2026 DeepSeek opened its MIT-licensed agent harness to developers as a developer preview. |
| API peak and off-peak pricing | August 16, 2026 Time-based rates began at 16:00 UTC. |
| DeepSeek-V4-Flash-Vision-Exp | August 21, 2026 Experimental multimodal API model with mixed text and image input. The same update added the Files API and Harness 0.1.1 support. |
DeepSeek Release History
| Release | Date and key change |
|---|---|
| DeepSeek-V4-Flash-Vision-Exp and Files API | August 21, 2026 Experimental multimodal API model that matches V4-Flash on text capabilities and accepts mixed text and image input through Chat Completions, Messages, and Responses. Images use up to 384 billed input tokens each at V4-Flash rates; the Files API supports reusable file_id inputs, and Harness 0.1.1 adds built-in support. |
| API peak and off-peak pricing | August 16, 2026 Time-based billing began at 16:00 UTC, with off-peak rates set at half the peak rates. |
| DeepSeek Harness v0.1 | August 13, 2026 MIT-licensed agent harness released in developer preview with a plugin-based architecture. |
| DeepSeek-V4-Pro-0813 API update | August 13, 2026 Current Pro model with a 1M context window, up to 384K output, and native Responses and Anthropic API support. |
| DeepSeek-V4-Flash API public beta | July 31, 2026 Post-trained Flash update for agent work, with native Responses API support and Codex integration. |
| DeepSeek-V4 Preview | April 24, 2026 Preview family with Pro and Flash models, dual reasoning modes, and a 1M context window. |
| DeepSeek-OCR-2 | January 27, 2026 Second document-understanding model based on visual causal flow. |
| DeepSeek-V3.2 | December 1, 2025 Added thinking during tool use. Open weights remain available after V4 replaced it on the official API. |
| DeepSeekMath-V2 | November 27, 2025 Mathematical reasoning model built around verification and proof-oriented training. |
| DeepSeek-OCR | October 20, 2025 Research model for document recognition and visual-text compression. |
| DeepSeek-V3.2-Exp | September 29, 2025 Experimental release that introduced DeepSeek Sparse Attention. |
| DeepSeek-V3.1-Terminus | September 22, 2025 V3.1 update for language consistency and agent performance. |
| DeepSeek-V3.1 | August 21, 2025 Combined thinking and non-thinking modes in one model and improved tool use. |
| DeepSeek-R1-0528 | May 28, 2025 Updated R1 with stronger reasoning, fewer hallucinations, JSON output, and function calling. |
| DeepSeek-Prover-V2 | April 30, 2025 Formal theorem-proving models released in 7B and 671B sizes. |
| DeepSeek-V3-0324 | March 24, 2025 Improved reasoning, coding, writing, and tool use. Model weights moved to the MIT License. |
| Janus-Pro | January 27, 2025 Updated unified multimodal models for image understanding and generation. |
| DeepSeek-R1 | January 20, 2025 Reasoning family with R1-Zero, R1, and six distilled models. |
| DeepSeek App | January 15, 2025 Official iOS and Android App, initially powered by V3. |
| DeepSeek-V3 | December 26, 2024 671B-parameter MoE model with 37B active parameters per token. |
| DeepSeek-VL2 | December 13, 2024 Second vision-language family using a mixture-of-experts design. |
| DeepSeek-R1-Lite-Preview | November 20, 2024 Hosted reasoning preview that preceded the open R1 release. |
| Janus | October 2024 Unified multimodal research family for image understanding and generation. |
| DeepSeek-V2.5 | September 5, 2024 Merged the general abilities of V2 Chat with the coding abilities of Coder-V2. |
| DeepSeek-Coder-V2 | June 17, 2024 MoE coding family with support for 338 programming languages and 128K context. |
| DeepSeek-Prover | May 2024 First DeepSeek research family for formal theorem proving in Lean. |
| DeepSeek-V2 | May 6, 2024 Introduced the general MoE architecture and Multi-head Latent Attention used by later flagships. |
| DeepSeek-VL | March 2024 First DeepSeek vision-language family. |
| DeepSeekMath | February 2024 Math-focused 7B models released as Base, Instruct, and RL checkpoints. |
| DeepSeekMoE | January 11, 2024 Introduced fine-grained expert routing and shared-expert isolation. |
| DeepSeek LLM | November 29, 2023 First general-purpose DeepSeek language models, released in 7B and 67B sizes. |
| DeepSeek Coder | November 2, 2023 First public DeepSeek model family, built for code generation and completion. |
Current DeepSeek Model Lineup
DeepSeek-V4-Pro-0813 is the current Pro model and is called through deepseek-v4-pro. It supports thinking and non-thinking modes within a one-million-token context window and can return up to 384K output tokens. The API supports JSON output, tool calls, the Responses API, the Anthropic API, chat prefix completion, and non-thinking FIM completion.
DeepSeek-V4-Flash-0731 is the current Flash version under the deepseek-v4-flash name. It has 284 billion total parameters and activates 13 billion for each token. DeepSeek re-post-trained the preview model for agent tasks without changing its architecture or size. Like Pro, it supports a 1M context window, dual reasoning modes, the Responses API, and the Anthropic API.
DeepSeek-V4-Flash-Vision-Exp is an experimental multimodal API model under the deepseek-v4-flash-vision-exp name. It matches V4-Flash on text capabilities while accepting mixed text and image input through Chat Completions, the Anthropic-compatible Messages API, and the Responses API. Images can be supplied as base64 data, external URLs, or Files API file IDs, and each image uses up to 384 billed input tokens at V4-Flash rates.
DeepSeek-V3.2 was the previous general flagship. Its open weights remain available, but V4 has replaced it on DeepSeek’s main API routes. V3.2 added reasoning during tool use and completed the architecture first tested in V3.2-Exp.
DeepSeek Harness and API Pricing
DeepSeek Harness v0.1, also called dsh, is an open-source agent harness built on the Cordis meta-framework. Models, tools, skills, sessions, sandboxes, filesystems, agent loops, orchestration, and the user interface are implemented as replaceable plugins. The August 13 release is a developer preview under the MIT License, and DeepSeek warns that compatibility-breaking changes will occur during development. Harness 0.1.1 added out-of-the-box support for V4-Flash-Vision-Exp on August 21, 2026.
DeepSeek’s API now uses peak and off-peak rates. Time-based billing began at 16:00 UTC on August 16, 2026, with off-peak prices set at half the peak rates. Peak hours are 01:00–04:00 and 06:00–10:00 UTC. For every one million tokens, V4-Flash will cost $0.014 for a cache-hit input, $0.44 for a cache-miss input, and $1.32 for output during peak hours. V4-Pro peak rates will be $0.044, $1.32, and $3.96 respectively. Images sent to V4-Flash-Vision-Exp use the V4-Flash rates after conversion to input tokens.
DeepSeek Timeline
2026: Vision, OCR-2, DeepSeek-V4, Harness, and API Changes
| Date | Release or Milestone and Why It Mattered |
|---|---|
| August 21, 2026 | DeepSeek-V4-Flash-Vision-Exp: DeepSeek added an experimental multimodal API model that matches V4-Flash on text capabilities and accepts mixed text and image input through Chat Completions, Messages, and Responses. Images use up to 384 billed input tokens each at V4-Flash rates. The Files API lets users reuse uploaded images through file_id, and Harness 0.1.1 added built-in support. |
| August 16, 2026 | API peak and off-peak pricing: DeepSeek’s time-based rates began at 16:00 UTC. Off-peak prices are half the peak rates, while peak rates for both V4 models are higher than the prices in effect before the change. |
| August 13, 2026 | DeepSeek Harness v0.1: DeepSeek released its MIT-licensed agent harness in developer preview. Its Cordis-based architecture implements models, tools, skills, sessions, sandboxes, filesystems, loops, orchestration, and the user interface as plugins that developers can replace or extend. |
| August 13, 2026 | DeepSeek-V4-Pro-0813 update: V4-Pro-0813 became the version served through deepseek-v4-pro. The model supports a 1M context window, up to 384K output, both reasoning modes, tool calls, and native Responses and Anthropic API formats. |
| July 31, 2026 | DeepSeek-V4-Flash public beta: DeepSeek released a re-post-trained version of Flash with stronger agent performance while retaining the preview model’s architecture and size. Native Responses API support made the model available in Codex. The update applied only to the Flash API; V4-Pro, the App, and web models were unchanged. |
| April 24, 2026 | DeepSeek-V4 Preview: DeepSeek released V4-Pro and V4-Flash through its web service, API, and open-weight repositories. Pro handles more demanding work, while Flash favors speed. Both have a 1M context window and selectable thinking modes. |
| January 27, 2026 | DeepSeek-OCR-2: The second OCR research release replaced the first model’s standard visual processing order with visual causal flow. The change was designed to compress long documents into more efficient visual representations. |
2025: R1, Hybrid Reasoning, Agents, and Specialist Models
| Date | Release or Milestone and Why It Mattered |
|---|---|
| December 1, 2025 | DeepSeek-V3.2: V3.2 arrived in the App, web service, API, and open-weight repositories. It could use tools in both thinking and non-thinking modes. A temporary V3.2-Speciale endpoint offered heavier reasoning until December 15. |
| November 27, 2025 | DeepSeekMath-V2: DeepSeek returned to specialist mathematical reasoning with a much larger model centered on verifiable generation. The release kept the separate math research line active after R1 brought advanced math into a general reasoning model. |
| October 20, 2025 | DeepSeek-OCR: This 3B document model represented text-heavy pages with compact visual tokens. It established OCR as a separate research branch rather than a feature inside the main chat model. |
| September 29, 2025 | DeepSeek-V3.2-Exp: The experimental model introduced DeepSeek Sparse Attention to reduce the cost of long-context processing. It served as the public architecture test for V3.2. |
| September 22, 2025 | DeepSeek-V3.1-Terminus: This corrective update reduced Chinese-English mixing and stray characters. It also improved code-agent and search-agent behavior without starting a new model generation. |
| August 21, 2025 | DeepSeek-V3.1: V3.1 put thinking and non-thinking behavior in one model and improved tool use for multi-step agent tasks. The V series was no longer limited to conventional chat and code completion. |
| May 28, 2025 | DeepSeek-R1-0528: The update improved mathematics, coding, and front-end generation. JSON output and function calling also made R1 easier to use in software integrations. |
| April 30, 2025 | DeepSeek-Prover-V2: DeepSeek released theorem-proving models in 7B and 671B sizes. The larger version brought the formal-proof branch onto the V3-scale architecture. |
| March 24, 2025 | DeepSeek-V3-0324: The update improved reasoning, coding, Chinese writing, search, and function calling. DeepSeek released the revised weights under the MIT License. |
| January 27, 2025 | Janus-Pro: Janus gained better image understanding and generation. It remained a downloadable research family, not a multimodal mode in the main DeepSeek App. |
| January 20, 2025 | DeepSeek-R1: DeepSeek released R1-Zero, R1, and six smaller distilled models. Its reinforcement-learning approach improved step-by-step reasoning and drew attention far beyond the developer community. |
| January 15, 2025 | DeepSeek App: The official App launched on iOS and Android with V3 underneath. R1 followed five days later as a separate model release. |
2024: MoE Research, V2, V3, and the First Reasoning Preview
| Date | Release or Milestone and Why It Mattered |
|---|---|
| December 26, 2024 | DeepSeek-V3: V3 expanded the company’s MoE design to 671 billion total parameters, with 37 billion active per token. It became the base for R1 and powered the DeepSeek App at launch. |
| December 13, 2024 | DeepSeek-VL2: The second vision-language family adopted an MoE design and came in Tiny, Small, and full-size variants. DeepSeek kept it separate from the text-only V3 family. |
| November 20, 2024 | DeepSeek-R1-Lite-Preview: DeepSeek put an early reasoning model on its web and API services without releasing the weights. The preview let users try the approach that later became R1. |
| October 2024 | Janus: Janus assigned different visual encoding paths to image understanding and generation but shared one language-model core. That split allowed one system to handle both jobs without forcing them through the same visual representation. |
| September 5, 2024 | DeepSeek-V2.5: DeepSeek combined V2 Chat and Coder-V2. General conversation and code work no longer needed separate official API model families. |
| June 17, 2024 | DeepSeek-Coder-V2: The second Coder generation adopted an MoE architecture, increased language coverage from 86 to 338 programming languages, and expanded context from 16K to 128K. |
| May 2024 | DeepSeek-Prover: The first Prover models targeted formal theorem proving in Lean. Unlike DeepSeekMath’s natural-language answers, a Lean proof assistant could check these proofs directly. |
| May 6, 2024 | DeepSeek-V2: V2 introduced Multi-head Latent Attention and a larger MoE design. Both efficiency ideas remained central to V3, R1, and later V-series models. |
| March 2024 | DeepSeek-VL: DeepSeek’s first vision-language models could interpret images, charts, and text within images. They opened a multimodal research line alongside the text models. |
| February 2024 | DeepSeekMath: DeepSeek published Base, Instruct, and reinforcement-learning versions of a 7B math model. Its Group Relative Policy Optimization method later became part of the R1 training work. |
| January 11, 2024 | DeepSeekMoE: The research model divided experts into smaller units and isolated shared knowledge in dedicated experts. DeepSeek carried this sparse design into its later flagship models. |
2023: DeepSeek Begins with Code and General Language Models
| Date | Release or Milestone and Why It Mattered |
|---|---|
| November 29, 2023 | DeepSeek LLM: DeepSeek released general-purpose 7B and 67B models in Base and Chat forms. They were the company’s first public models for language tasks beyond programming. |
| November 2, 2023 | DeepSeek Coder: DeepSeek’s first public model family handled code completion, generation, and repository-level context. Coding was the company’s starting point, not a later addition to a chat model. |
| 2023 | Company founding: Liang Wenfeng founded DeepSeek in Hangzhou with financial backing from High-Flyer. The new company pursued artificial general intelligence research rather than extending High-Flyer’s trading business. |
DeepSeek Model Families Explained
V Series
The V series is DeepSeek’s main general-purpose model line. V2 introduced the MoE and attention architecture that shaped later releases. V2.5 combined general and coding abilities, V3 increased the scale, V3.1 merged thinking with regular response modes, and V3.2 brought reasoning into tool use. V4 divided the main line into Pro and Flash models; Flash became the first to receive an official API release, and V4-Flash-Vision-Exp later added experimental image input to that API family.
R Series
The R series focuses on reasoning. R1-Lite-Preview was the first hosted preview. R1-Zero tested reinforcement learning without supervised fine-tuning as the first post-training stage, while R1 added cold-start data to improve readability and consistency. The six distilled R1 models transferred parts of that reasoning behavior into smaller Qwen- and Llama-based models.
Coder, Math, and Prover
DeepSeek Coder and Coder-V2 were built for programming. DeepSeekMath concentrated on competition-style mathematical problems expressed in natural language. DeepSeek-Prover targeted formal proofs that a Lean proof assistant could check. These branches influenced the general models even when their names disappeared from the main API.
VL, Janus, and OCR
DeepSeek-VL and VL2 handle vision-language understanding. Janus combines image understanding with image generation, but uses different visual encoding paths for the two jobs. DeepSeek-OCR treats document images as compressed carriers of text and layout. The newer DeepSeek-V4-Flash-Vision-Exp adds an experimental multimodal API branch that accepts images with text through Chat Completions, Messages, and Responses. It remains a separate vision variant alongside the general V4-Pro and V4-Flash endpoints.
What the Model Suffixes Mean
| Suffix | Meaning |
|---|---|
| Base | A pretrained checkpoint before instruction tuning for conversational use. |
| Chat or Instruct | A model tuned to follow instructions and respond in a conversational format. |
| Distill | A smaller model trained with outputs or reasoning traces from a larger model. |
| Lite | A smaller model or preview intended to use fewer computing resources. |
| Preview | An early public release that is not presented as the final version. |
| Exp | An experimental release used to test a new architecture or serving method. |
| Terminus | The final corrective update in the V3.1 line before the V3.2 experimental branch. |
Five Turning Points in DeepSeek’s History
1. The First Releases Were Specialist Models
DeepSeek did not begin with a consumer chatbot. It released a code model first, then a general language model, followed by separate research families for MoE routing, mathematics, and vision. Each branch tested an idea that could later move into a larger general model.
The specialist work also fed later releases. DeepSeekMath explored reinforcement-learning methods used in R1. DeepSeekMoE supplied the sparse expert design behind V2 and V3. Coder and Math data helped train Coder-V2 and the general models that followed.
2. V2 Established the Efficiency Architecture
V2 was the first general DeepSeek model built around the architecture now associated with the company. Its mixture-of-experts system activated only part of the full parameter set for each token. Multi-head Latent Attention reduced the memory required for the key-value cache during inference.
Those choices carried into later releases. V3 used a much larger MoE design, R1 inherited the V3 base architecture, and V4 continued the effort to lower the cost of long-context inference. V2 is the architectural starting point for the later flagships.
3. V3 Became the General Foundation
V3 expanded the MoE design to 671 billion total parameters while activating 37 billion for each token. It handled general writing, coding, mathematics, and other text tasks through one model. The DeepSeek App used V3 at launch, and the official deepseek-chat API identifier pointed to V3 at the time.
V3 also became the base for R1. R1 was not a separate architecture built from the ground up. It applied a reasoning-focused post-training process to the capabilities and efficiency of the V3 foundation.
4. R1 Took DeepSeek Beyond Its Developer Audience
R1 made DeepSeek widely known in January 2025. The release combined a large reasoning model, an experimental R1-Zero checkpoint, and six smaller distilled models. Developers could inspect and run the weights, while ordinary users could access reasoning through the DeepSeek service.
The App had launched five days earlier with V3. R1 expanded what users could do inside the existing product, but it was not the App’s launch model. The later R1-0528 update added JSON output and function calling for software integrations.
5. V3.1 and V4 Merged Reasoning with Agent Work
V3.1 reduced the separation between a normal chat model and a reasoning model. One model could answer in thinking or non-thinking mode, and the API identifiers selected the mode. V3.2 extended that design by allowing the model to reason while using tools.
V4 kept both modes and gave Pro and Flash a 1M context window. The July 2026 Flash update strengthened multi-step agent work and added native Responses API support, which allowed Codex clients to call the model through DeepSeek’s API. V4-Flash-Vision-Exp extended that API path to mixed text-and-image requests on August 21, 2026.
Read More: How to use DeepSeek v4 Pro & Flash in Codex and Cut API costs
DeepSeek Timeline FAQs
When was DeepSeek founded?
DeepSeek was founded in Hangzhou, China, in 2023 by Liang Wenfeng. High-Flyer, the quantitative investment firm he co-founded, provided financial backing. No single founding day is consistently documented because the research lab’s creation and its later corporate separation occurred at different stages.
What was the first DeepSeek model?
DeepSeek Coder was the first public DeepSeek model family. It was released on November 2, 2023, with models for code completion and instruction-based programming tasks. DeepSeek LLM followed on November 29 as the company’s first general-purpose language model family.
When was DeepSeek-R1 released?
DeepSeek-R1 was released on January 20, 2025. The release included DeepSeek-R1-Zero, the main R1 model, and six distilled models based on Qwen and Llama architectures. DeepSeek updated the reasoning family with R1-0528 on May 28, 2025.
Did the DeepSeek App launch with R1?
No. The official DeepSeek App launched on January 15, 2025, and was initially powered by DeepSeek-V3. R1 arrived on January 20 and was then made available through DeepSeek’s web product, App, and API. The two events happened in the same week but were separate releases.
What is the latest DeepSeek model?
DeepSeek-V4-Flash-Vision-Exp is the newest model announced as of August 21, 2026. It is an experimental multimodal API model served through deepseek-v4-flash-vision-exp, matches V4-Flash on text capabilities, and accepts image input through Chat Completions, Messages, and Responses. DeepSeek-V4-Pro-0813 remains the current Pro model, while DeepSeek-V4-Flash-0731 remains the general Flash option.
Does DeepSeek-V4-Flash-Vision-Exp support images?
Yes. The experimental model accepts JPEG, PNG, GIF, and WebP images alongside text. You can provide an image as base64 data, an external URL, or a Files API file_id through the Chat Completions, Messages, or Responses API. Images are converted to input tokens for billing, with a maximum of 384 tokens per image at V4-Flash rates.
What is the difference between DeepSeek V models and R models?
The V series is DeepSeek’s main general-purpose line. It covers ordinary chat, writing, coding, tool use, and other text tasks. The R series was created for deliberate reasoning, especially mathematics, coding, and multi-step problems. Starting with V3.1, the V series also gained thinking modes, so the boundary is less rigid than it was when R1 launched.
Is DeepSeek open source?
Most DeepSeek models have downloadable weights, technical reports, and permissive licenses. R1 and several recent V-series releases use the MIT License.
What do deepseek-chat and deepseek-reasoner mean?
deepseek-chat and deepseek-reasoner are legacy API compatibility identifiers whose underlying models changed over time. They most recently selected V4-Flash’s non-thinking and thinking modes.
Related Resources
- OpenAI & ChatGPT Timeline
- Anthropic Claude Timeline
- Google AI & Gemini Timeline
- Meta AI Timeline
- Reasonix: Free DeepSeek Coding Agent with Cache-First Cost Control
- CodeWhale (DeepSeek TUI): Free, Open-source Claude Code Alternative
- Open-Source Browser Automation & AI Assistant – DeepSeek Cowork
- DeepSeek Harness: Open-Source Plugin-Based AI Agent Harness
- Best DeepSeek Harness Plugins for DSH










