← Glossary Model
DeepSeek V4-Flash
The smaller, cheaper tier of DeepSeek's V4 model family, priced around $0.14 per million input tokens with a one million token context window.
V4-Flash sits below the main V4 model in DeepSeek’s lineup: roughly $0.14 per million input tokens and $0.28 per million output, against about $0.44 and $0.87 for the larger tier. Both handle a context window of a million tokens, which is very roughly how much text the model can keep in mind at once.
Like the rest of DeepSeek’s releases it is published with open weights, so you can download and run it yourself rather than only calling an API. A vision-capable experimental version arrived in September 2026.
Mentioned in
-
DeepSeek Opens Up a 305 Billion Parameter Model That Can Finally See
-
DeepSeek's Cheap Model Can Now See, and It Edges Past Opus 4.8 on Two Visual Tests
-
Someone got DeepSeek V4 Flash running in production on a single AMD card and published every patch it took
-
DeepSeek updated its cheap model and it now runs neck and neck with OpenAI's cheap model
-
DeepSeek V4 goes fully stable, and the old models switch off today