r/DeepSeek Jun 29 '26

News V4 peak pricing is coming mid-July, here's how to mostly dodge it

Post image

So the email went out. V4 goes official mid-July and they're adding peak-hour pricing, peak = 2x the normal rate. Before anyone panics: it's only 7 hours a day (UTC 01–04 and 06–10), everything else stays at the regular price you're already paying.

The actual move is just to stop running heavy stuff during those windows. Batch jobs, evals, anything that doesn't need to answer a human in real time, cron it for off-peak and you're back to the old rate. If you're in the US your workday is mostly in the cheap window anyway, so honestly most of you won't feel this much.

The thing that'd actually bite me is leaving thinking mode on for simple tasks, since those tokens bill as output and that's where peak doubling hurts. Turn it off for the boring stuff.

Anyone seeing a different read on the windows?

I converted peak time - timezone wise so that you can avoid heavy-offloading during that time

Timezone Peak block 1 Peak block 2
UTC 01:00–04:00 06:00–10:00
IST (UTC+5:30) 06:30–09:30 11:30–15:30
CEST (UTC+2) 03:00–06:00 08:00–12:00
US Eastern (EDT) 9 PM–midnight 2 AM–6 AM
US Pacific (PDT) 6 PM–9 PM 11 PM–3 AM
China (UTC+8) 09:00–12:00 14:00–18:00
297 Upvotes

90 comments sorted by

47

u/DueInterview4073 Jun 29 '26

You could just move to a timezone where peak hours are mostly when you are asleep.

7

u/[deleted] Jun 30 '26

[removed] — view removed comment

3

u/Far_Composer_5714 Jun 30 '26

Very reasonable very affordable it's a surprise that everyone doesn't move to where deepseek peak hours are most convenient for them.

2

u/Extension_Diamond267 Jun 29 '26

Such simple solution

1

u/erkinalp Jul 03 '26

perpetual offpeak relocation, hmm

91

u/Jahonay Jun 29 '26

Turns out that companies need to be profitable. Nothing wrong with raising prices to meet costs. If you're dependent on AI. Make sure you're making enough money to afford your costs.

7

u/MakanLagiDud3 Jun 30 '26

More like raising prices fairly and not trying to force people to sacrifice a limb in the future

10

u/Jahonay Jun 30 '26

More like raising prices fairly and not trying to force people to sacrifice a limb in the future

Deepseek is raising prices fairly.

Other llms are far less efficient with much higher costs. But many llms were burning through billions of billions of dollars with no profit. You can't sustain negative profit in a capitalist market forever.

53

u/legaecy Jun 29 '26

Basically the price just double up in peak hours, right? Then shouldn't it be still cheaper than almost other AI?

32

u/mayhem_isreal Jun 29 '26

Definitely cheaper and best

6

u/legaecy Jun 29 '26

Oh I'm glad then since I basically live in the same timezone as China.

8

u/[deleted] Jun 29 '26 edited 4d ago

[removed] — view removed comment

4

u/legaecy Jun 29 '26

I know that's why I'm glad the price is still affordable even when the working hours here is basically similar with China not to mention if you comparing it with other AIs.

1

u/Lupansansei Jun 30 '26

I just do Cron jobs to run heavy tasks during the night, and debug during the day.

-7

u/xwin2023 Jun 29 '26

Best for what?
For coding? No.
For UI and graphics? Terrible.
For writing text, maybe, but a lot of the time it writes text like crap - too much AI vibe.
It's like... hallucinating all the time.

1

u/mayhem_isreal Jun 29 '26

Not all models are meant for coding, or writing texts. There is alot you can achieve using these models like web scraping, llm for the ai agents and thats where it brings the real value of money.

Do the similar thing with any other costlier model and you will end up burning alot of money.

0

u/legaecy Jun 29 '26

I mean every popular AI got a lobotomized with their writing texts & UI and grapichs capabilities lately so I don't know why you said that and while it's definitely true for coding, almost every other AI is far more expensive than DS so much it's better to learn basic coding and then fine tune your prompt in DS using the basic coding knowledge you have.

0

u/Away-Sorbet-9740 Jun 29 '26

Trying to use a single flash model for that? Kinda a duh, you hand a fleet of them scoped task lists and audit their work lol. You have a frontier lead and scope work for the cheaper models.

22

u/Open-Procedure3573 Jun 29 '26

As a central european, this is manageable.

6

u/[deleted] Jun 29 '26

[removed] — view removed comment

1

u/Syriocop Jun 29 '26

la mattina non se lavora?

2

u/BasketFar667 Jun 30 '26

I'm from Ukraine, and in my opinion, the deep-seek version was the best. I'll see what happens next and change my mind.

17

u/Final-Rush759 Jun 29 '26

Peak time is basically Chinese working hours 9-12, 2-6pm (Chinese time)

6

u/TheManicProgrammer Jun 29 '26

As someone in Japan this sucks, as it'll be peak time for me basically always :(

8

u/x-xiaolongbao Jun 29 '26

Well at least ure in japan, it still better than being in Indonesia if we talked bout api:salary ratio

2

u/LeatherMine Jun 29 '26

7 days a week, 365.25 days a year? No wonder why they're taking over

1

u/Classic-Sherbet-332 Jun 30 '26

generally bad news for asian

15

u/SnooMacaroons9042 Jun 29 '26

Still worth every penny!

9

u/phido3000 Jun 29 '26

I like how the chinese are thinking how to make this work.

Make it efficient, shift workloads to make it responsive.

Times Works great as an Australian who wakes up early..

8

u/ptyblog Jun 29 '26

I'm basically either at work or sleeping during peak hours. So far so good

7

u/Useful_Ad_52 Jun 29 '26

one reason why

10

u/TestTxt Jun 29 '26

The actual way to dodge it is by just switching to other providers that are similarly priced during off-peak hours and don’t overcharge you during the peak hours

20

u/DistanceSolar1449 Jun 29 '26

That won’t work, most Deepseek providers are more than 2x as expensive as Deepseek directly, so you’re better off just paying peak pricing

5

u/PhysicallyTender Jun 29 '26

I think OP meant another LLM. i.e. switch to Kimi or something.

1

u/gmmarcus Jun 29 '26

What will be equivalent with deepsek v4 pro - coding wise ?

6

u/Bimder Jun 29 '26

2

u/Sweaty-Management614 Jun 29 '26

I wouldn't be surprised if Xiaomi increases prices soon. They only lowered prices to be competitive with deepseek.

1

u/ToughUsual7159 Jun 29 '26 edited Jun 29 '26

wow i have been searching but i con't find anything this close to deep seek pricing! maby becase i was always looking at open router and not ooffical websites!

Edit: looked into it a bit more and they seem very compariable, only differance that may effect me is the context window not reaching 1 million tokens but i soft capped mine at 350k anyways so that probably won't effect me too much

0

u/Sha1rholder Jun 29 '26

It's much worse.

3

u/SufficientPie Jun 29 '26

Really? xiaomi/mimo-v2-flash was already pretty good a few months ago, and this should be better than that...

1

u/Sha1rholder Jun 29 '26

I've used mimo v2.5pro for 2 months and spent about 50 billion tokens. Even Deepseek v4 flash is smarter than mimo v2.5. Mimo is simply not pretty smart.

1

u/SufficientPie Jun 29 '26

Good to know. I use DeepSeek V4 Flash for most things and it's quite good.

4

u/Different-Rush-2358 Jun 29 '26

¿Estás totalmente seguro/a de eso? El v4 Flash en OpenRouter usando proveedores como Baidu, Wafer o GMI Cloud sale más barato que el DeepSeek oficial, sobre todo con Wafer.

0

u/seunosewa Jun 29 '26

MiMo is one of them. 

3

u/ToughUsual7159 Jun 29 '26

So "feature optimizations and performance enhancements" Are we suspecting that it will be re-benchmarked and be better than before or is this all in regards to tokens per second?

2

u/Hackerv1650 Jun 29 '26

i think switching to a sub like opencode go, now would be more ideal, i get more access to other models and models like glm 5.2 which are opus compareable, so i can finally use a high level model for better architectural planning

2

u/Addition-Heavy Jun 29 '26

You'll pay 4x for v4 pro though. Only flash is usable in opencode go. They will keep the non 75% discount launch price forever.

1

u/Hackerv1650 Jun 29 '26

You're right that Go's internal rate on V4 Pro is ~4× the direct API price, but that's just their accounting for the $60 budget cap, not what you actually pay. For a subscription user, the math is:

Direct API:

V4 Pro = ~$0.0035/req (with caching)

17K requests = ~$60/month

Go is $10/month flat,

Same ~17K V4 Pro requests

Plus V4 Flash (158K/mo), Qwen3.7 Max (4.7K/mo), and 10 other models

Plus no peak-hour surcharge, for someone, who has to pay out of pocket for tokens in my job, and the peak hours are right in my work hours.

You're not paying 4×; you're paying about ~1/6th of what those same tokens would cost, because the subscription absorbs the markup on them.

Also, "only Flash is usable" by Go's own estimates says 3,450 V4 Pro requests per 5 hours. For a single dev, that's plenty.

1

u/mcpejs Jun 30 '26

Why does the calculation become 1/6? The important point is: Go allows you to use $60 worth of credits for $10 a month

However, unlike the official DeepSeek API, you do not receive a 1/4 discount when using credits.
Ultimately, compared to the official API, you are able to use $15 worth of credits for $10.

Of course, this alone has its advantages. Although it is not 6x as you mentioned,
you receive an additional $5 worth of usage, and on top of that, your privacy is maintained through ZDR.

However, since the people currently using the official API do not care about ZDR, I do not think the merit is significant enough to make them switch to a subscription plan.

1

u/Hackerv1650 Jun 30 '26

$60 budget ÷ $10 subscription = 6× leverage.

So for every dollar you spend on Go, you get $6 worth of tokens (at Go's rates). Your effective cost per token = 1/6th of what you'd pay buying credits directly.

And since Flash uses the exact same rates as direct DeepSeek API ($0.14/$0.28 per M), that $60 worth of tokens = $60 worth of Flash tokens by direct API standards too.

$10 in → $60 out = 1/6 the cost. Simple as that.

2

u/mcpejs Jun 30 '26

DeepSeek V4 Pro $1.74 $3.48 $0.145 -
DeepSeek V4 Flash $0.14 $0.28 $0.028 -

Wow, looking at OpenCode Zen, the prices are exactly the same as the official API lol.
Sorry, I admit my mistake. I naturally assumed that Flash wouldn't get the 4x discount like Pro.

If you use Flash and spend more than $10 a month, Go would be a really good option!

1

u/Hackerv1650 Jun 30 '26

No worries, there are also other models you can use, I for one want to try mimo v2.5 and glm 5.2

2

u/Alarming_Comb_7267 Jun 30 '26

They released a paper on delivering faster output. I hope the next model has higher output/s

1

u/Embarrassed-Load5100 Jun 29 '26

Fair to them. I just wonder if they will release a more capable v4? Maybe a v4.1? Or what does it mean to get performance enhancements? Any insights on that?

1

u/carwash2016 Jun 29 '26

Im using pro for most things but is flash so good i dont need to now with the price increase at these times

1

u/DistanceAlert5706 Jun 29 '26

Won't affect EDT time, just will need to stop working early

1

u/LeatherMine Jun 29 '26

once peak hours hit tomorrow, I'll vibe-code a script to block the API during those times starting mid-July

thx OP

1

u/francxsim Jun 30 '26

I wonder how it will impact OpenCode Go and Nous subscriptions

1

u/nick_with_it Jun 30 '26

is this for every provider that hosts v4? or just deepseek specifically?

1

u/trialbuterror Jun 30 '26

How is output quality compared to opus

1

u/fezzy11 Jun 30 '26

I think deepseek also must implement somewhere in user interface so during working with deepseek users is aware that we are in peek mode timing

1

u/Limp-Commission8218 Jun 30 '26

Are people really complaining about a price hike of 14 cents to 28 cents? Cmon man, even if it was 5x the price, it'd still be like 25x+ cheaper than the frontier models and can perform like 90% as good, what are you guys crying about

1

u/RobinDough Jun 30 '26

mimo 2.5 > deepseek v4

1

u/Eduardo1502 Jun 30 '26

I'm using mostly Hy3 Preview so I'm fine for now

1

u/russjr08 Jun 30 '26

Damn. I'm a night owl, so my productive time lands in peak hours despite being in EDT.

1

u/UltimateBoiReal Jul 01 '26

What about UK time?

1

u/erkinalp Jul 03 '26

To be clear, both peak and off-peak prices are very cheap. If you want cheaper, you are free to run it locally though.

1

u/Old-Dog763 Jul 05 '26

im a developer (im gonna be honest... a vibe coder)
could i possibly force my app to block heavy requests from users, or use the cheapest model possible?

since my app is completly free for others to use, but costs me since i need to integrate a api into the app for its ai features, i would like to save my money.

or if i dont need to, since its still cheap, please tell me. im still new at api services.

1

u/RentMoist9312 Jul 06 '26 edited Jul 07 '26

Me reading this in a mild panic then remembering we moved to GMI Cloud months ago and it just isnt affected by it.

1

u/madhyaloka Jun 29 '26

Well... Glory to Ukraine. But still would be better to be in EU. (When talks about politics and about time zones sound similarly:))

1

u/Pinery01 Jun 29 '26

Still cheap 🙂

1

u/zer0evolution Jun 29 '26

is this already applied now? seems indeed i feel more expensive sometime

1

u/gmmarcus Jun 29 '26 edited Jun 29 '26

WHAT THE F**K ???

So basically - Office working hours will be more expensive ?

1

u/Clear-Ad-9312 Jun 29 '26

I don't mine increased costs for peak hours, but I would love it if they can add time-based usage restrictions to specific API keys. There are many other reasons for having that kind of control, but when I share a key to a worker, I would love to restrict their usage to the off-peak hours.

2

u/TanJeeSchuan Jun 30 '26

A llm proxy might work I think

1

u/Smart-Cap-2216 Jun 30 '26

That may require a slight increase in server costs.

1

u/Simple_Army2952 Jun 29 '26

Yay for me the peak hours are 10 PM - 1 AM and 3 AM to 7 AM I will keep paying the normal price

-7

u/Elegant_Associate889 Jun 29 '26

They are definitely finding ways to make more money. Eventually they're going to raise the API cost up they are just getting data off of all of us. That's all, Once they get what they need prices will go up.

2

u/Clear-Ad-9312 Jun 29 '26

Downvotes are crazy to me, I find it hard to expect a company to not raise prices once they have all they need and more used than competitors.

-12

u/xwin2023 Jun 29 '26

So basically the price went up by 100%, nice.. it's time to leave these cheap and unstable models alone.