Rendered at 12:22:08 GMT+0000 (Coordinated Universal Time) with Cloudflare Workers.
j1elo 1 minutes ago [-]
So many changes in so little time, that it all makes no sense. Continuous churning. Reminds me of the experience of trying to be on top of the dependencies in a medium-large JS project.
I am a person that buys into a tool or a process and expects it to be part of the life with no major changes through the years (or as long as the need exists). But AI? You buy into something today, not 2 weeks have passed and there's already a large "update" introduced to the conditions or the optimal usage patterns you should be adopting.
alexpotato 8 minutes ago [-]
I'm no expert in pricing economics but once peak/off-peak pricing arrives, it seems like tokens are going to be like electricity or long distance phone minutes where it just becomes a commodity/race to the bottom.
garrickvanburen 2 minutes ago [-]
Yes. I focus on pricing software and I’m a bit baffled why frontier models are pushing tokens.
It’s a race to the bottom, and the bottom is unlimited use for a flat monthly rate.
Granular pricing (tokens, minutes, etc) is pretty anti-customer generates less revenue than customer value-based subscriptions (why SaaS is such a good business model)
progval 2 hours ago [-]
Interesting to see that peak hours are work hours in China, night in the US and Europe, and also morning in Europe. So Deepseek's customers are mostly domestic.
thecopy 32 minutes ago [-]
Peak Hours: 01:00–04:00 and 06:00–10:00 UTC
For European and US customers this is effectively 2x increase. I think i wll keep using both Flash and Pro as before.
EDIT: Misread numbers to believe off-peak kept old prices
jLaForest 19 minutes ago [-]
~200% increase is marginal to you?
nchmy 2 minutes ago [-]
200% increase over practically free is still practically free
Hamuko 2 hours ago [-]
Not that surprised about it. Personally I've seen companies really just go all-in on a single provider, and that has usually been Anthropic. I don't think we're allowed to run Chinese models even locally.
cheesecakegood 1 hours ago [-]
Also 6-9pm Pacific I think is (coincidentally) peak so it hits the ‘after work hobbyists’ still, which is I suspect is their current main audience.
r00t- 33 minutes ago [-]
That's a bit obvious, isn't it?
roenxi 46 minutes ago [-]
This is somewhat funny when you realise the data centres are now going to start a process that looks very so slightly like daydreaming. Depending on the time of day they're going to be thinking about different things in a cyclic manner. They're going to be doing things like finishing a hard days work then kicking back to think about tricky math problems.
ssk42 31 minutes ago [-]
That’s what my KimiClaw has literally been doing
hopfenspergerj 26 minutes ago [-]
Does the API response include a "service tier" response to indicate whether you paid peak/off-peak for a given request? I like to compute cost for each request, and save it with my results.
alkonaut 2 hours ago [-]
There is no relative/percentage increases noted (understandably). Just because i'm lazy: roughly how much more expensive is it to work with v4 flash and v4 pro through the API, compared to before the price increases? Is it 2x, 5x, 10x higher?
zupa-hu 2 hours ago [-]
# Flash, off-peak
cache-hit 2.5x
cache-miss 1.57x
out 2.36x
# Flash, peak
cache-hit 5x
cache-miss 3.14x
out 4.71x
Edit: fixed the numbers and formatting
embedding-shape 1 hours ago [-]
Someone made a comparison yesterday, including relative increases, and GPT-5.6 Luna, then later someone also added more OpenAI, Anthropic, K3 and GLM 5.2: https://news.ycombinator.com/item?id=49286679
Already outdated though I think, as GLM 5.3 is latest now :)
floppyd 2 hours ago [-]
About 2x-2.5x off-peak for Flash, 2x-4x I'd say for Pro (x6 on cache in, the biggest increase throughout the board). And twice as much in peak hours.
2 hours ago [-]
acrush 4 minutes ago [-]
My friend told me the price was higher than the last version. Is that true?
With proprietary labs lowering their prices and Deepseek raising theirs over time, wouldn't it possible to extrapolate a graph to look at where the terminal frontier-model million-token-cost asymptotes to?
mateenah 46 minutes ago [-]
This is good for other competitors I guess. People rarely calculate the bump in price but the fact that price is increasing might bring them to other vendors.
poly2it 2 hours ago [-]
That's a hefty increase. Flash pricing during peak is now 1.32/M out, compared to the current 0.28/M, which in turn is a quite a bit above the cheapest provider at 0.16/M.
I am a person that buys into a tool or a process and expects it to be part of the life with no major changes through the years (or as long as the need exists). But AI? You buy into something today, not 2 weeks have passed and there's already a large "update" introduced to the conditions or the optimal usage patterns you should be adopting.
It’s a race to the bottom, and the bottom is unlimited use for a flat monthly rate.
Granular pricing (tokens, minutes, etc) is pretty anti-customer generates less revenue than customer value-based subscriptions (why SaaS is such a good business model)
For European and US customers this is effectively 2x increase. I think i wll keep using both Flash and Pro as before.
EDIT: Misread numbers to believe off-peak kept old prices
Already outdated though I think, as GLM 5.3 is latest now :)
https://news.ycombinator.com/item?id=49287881
https://news.ycombinator.com/item?id=49285160
https://openrouter.ai/deepseek/deepseek-v4-flash-0731#provid...
The big Covid migrations (startup prople migrating to the countryside),
Will we see the big AI migrations (people travelling to where AI is the cheapest)?
DeepSeek-V4-Flash (off-peak, x2 for peak)
* Cache Hit $0.007 (x2.5)
* Cache Miss $0.22 (x1.5)
* Output $0.66 (x2.25)
DeepSeek-V4-Pro (off-peak, x2 for peak)
* Cache Hit $0.022 (x6)
* Cache Miss $0.66 (x1.5)
* Output $1.98 (x2.25)
Peak Hours: 01:00–04:00 and 06:00–10:00 UTC
Effective from: 16:00, August 16, 2026 (UTC)
It’s still cheaper than everybody else.
https://openrouter.ai/deepseek/deepseek-v4-flash#providers
- but what matter - is cache hit
even now deepseek's off-peak hours for cache hit (0.007) is lower than other providers (~0.01)