DeepSeek was hugely underpricing cache hit pricing before and even after this increase they're still cheaper on that metric than every other provider I'm aware of, but it will put an end to those "I used 1 billion tokens and spent $4" reports.
Model | | Input | Output | Cache Read
------------------------------------|----------------|-----------------------------|---------------------|-----------------
DeepSeek-V4-Flash | Prev | $ 0.14 | $ 0.28 | $ 0.0028
| Off-Peak | $ 0.22 (1.6x) | $ 0.66 (2.4x) | $ 0.007 (2.5x)
| Peak | $ 0.44 (3.1x) | $ 1.32 (4.7x) | $ 0.014 (5.0x)
DeepSeek-V4-Pro | Prev | $ 0.435 | $ 0.87 | $ 0.003625
| Off-Peak | $ 0.66 (1.5x) | $ 1.98 (2.3x) | $ 0.022 (6.1x)
| Peak | $ 1.32 (3.0x) | $ 3.96 (4.6x) | $ 0.044 (12.1x)
gpt-5.6-luna: $0.20 / $1.20 / $0.02 / $0.25 (In / Out / Cache Read / Cache Write)EDIT: formatting
unified101•50m ago
Flavius•44m ago
m101•26m ago
cmrdporcupine•15m ago
But DeepSeek v4 Pro is a far more capable model and still cheaper than anything that it competes with, from what I can see.