CourionAI
EN
Newsletter
← Glossary Model

DeepSeek V4-Flash

The smaller, cheaper tier of DeepSeek's V4 model family, priced around $0.14 per million input tokens with a one million token context window.

V4-Flash sits below the main V4 model in DeepSeek’s lineup: roughly $0.14 per million input tokens and $0.28 per million output, against about $0.44 and $0.87 for the larger tier. Both handle a context window of a million tokens, which is very roughly how much text the model can keep in mind at once.

Like the rest of DeepSeek’s releases it is published with open weights, so you can download and run it yourself rather than only calling an API. A vision-capable experimental version arrived in September 2026.