DeepSeek v4.1 Flash is here with a new price

The new DeepSeek v4.1 Flash model is here and the two most important things about the model is that it has native vision capabilities, and it's significantly cheaper after the price hike last month. I have updated the pricing on the DeepSeek peak/off-peak clock page already.

DeepSeek v4.1 Flash announcement

Another interesting thing is, as you see above, the new v4.1 Flash has surpassed the existing v4 Pro on all metrics so from now on all requests to the Pro model will be routed to the new Flash model.

Interesting!

About the new price, here's the price table and comparison with the old pricing (all prices in per million tokens):

Metric Off-peak, old Off-peak, new Peak, old Peak, new Reduction
Input, cache hit $0.007 $0.003 $0.014 $0.006 57.1%
Input, cache miss $0.22 $0.15 $0.44 $0.30 31.8%
Output $0.66 $0.60 $1.32 $1.20 9.1%

I used the model a lot yesterday, and it's visible that it's significantly better than other models, and it just has a better taste. I tested it for writing, and it already writes better than other DeepSeek models and even more natural than Claude or GPT models.

Good job, DeepSeek! And thank you.