DeepSeek planning to significantly raise prices
Posted by miroljub 15 hours ago
Comments
Comment by Frannky 15 hours ago
Any suggestions for a better configuration? I mostly need Opus 4.6 + Claude Code alternative. I don't need Fable level capabilities, I add one new feature at a time and approve the code before shipping it. Then test and open a PR.
Usually both Fable and Opus suggest dumb ideas but are good implementers once I tweak the idea, approve the code, and add unit and production tests.
I'm OK spending max $100/month on APIs, ideally with Zero data retention. I only need a few hours a day of coding. I don't want agents running all the time; I figure I can stay on top to each feature and wrap my head around the product and new suggestions as long as I don't build too much at once.
I'm still on a Claude Max $100 plan, but it's barely usable anymore—one call and I hit 20–30% of the 5h window on Opus 4.8. Opus 5 seems tuned to make messes, and Fable burns tokens for a level of capability I don't actually need.
Comment by anentropic 14 hours ago
Comment by Frannky 13 hours ago
It's similar to the situation where you ask whether you should walk or take the car to go refill the car with gas. I feel LLMs have a linearity embedded in them that prevents them from finding non-linear, smart solutions, especially when data is scarce. They can probably get it done, but with far more complex solutions, which increases the risk of ending up with a crazy complex codebase when it could have been much simpler and more elegant.
The other situation is when I spot a problem or a feature change that's needed after experiencing the product, and I ask how to change the codebase, it suggests something, I say "why not this other thing," and then we do the other thing.
Comment by conception 8 hours ago
Comment by throwaw12 15 hours ago
Comment by alhimik45 11 hours ago
Regarding Pi my position is that it is brilliant piece of software if you don't need extensions - give the model bash tool and let it do all it wants through it or use Pi as SDK for your own advanced harness or smth similar.
But Pi with extensions has two problems. First one: rather often they don't play well together. For example, if you want some adjustments A and B for the same tool and there are two extensions which do A and B, they will likely not work as expected when installed simultaneously. You could say that it can be solved by adjusting extensions or just generating your own - yes and it is the second problem. Like any piece of code you own and use, you have to maintain it. Bug here, incompatibility there and voila - you spend your precious time to work on harness instead of doing your job. Plus remember that vibecoders are not very responsible people, so Pi extensions registry is flooded by "use Pi to customize Pi" buggy one shot extensions.
With carefully developed set of extensions Pi would be better, like properly configured Arch Linux could be better than Linux Mint in the hands of power user. But considering how fast things are changing in this sphere, seems it is more optimal to take more bloated harness - with unneeded tools, too big prompts, etc - which will be effective on 90%, but do the actual job with that harness right now.
Comment by DrammBA 14 hours ago
Comment by actsasbuffoon 13 hours ago
Comment by DrammBA 10 hours ago
Comment by Frannky 13 hours ago
Comment by yojo 14 hours ago
Not sure about the OAi Pro plan, doesn’t look like the 80% Luna price slash made its way into the quota system.
You could also try tuning down the effort level on Opus. It makes a huge difference in token consumption and you might be able to get away with lower than you’ve set
Comment by ignoramous 14 hours ago
MiniMax's "token plan" ($20/mo for 1.7b tokens) is cost competitive. MiniMax M3 is equally good, if not better than DeepSeek v4, at coding: https://platform.minimax.io/subscribe/token-plan?tab=individ...
If you prefer pay-as-you-go, then Xiaomi MiMo is the only other provider with comparable models (MiMo v2.5 & Pro) that matches DeepSeek's current API rates for input/output/cache: https://mimo.mi.com/docs/price/pay-as-you-go
Meanwhile, Meta is running a 10x discount on Muse Spark 1.2 (Grok 4.5 / Sonnet 5 level model), if you opt-in to data sharing: https://dev.meta.ai/docs/getting-started/pricing-rate-limits
> I'm OK spending max $100/month on APIs, ideally with Zero data retention.
In that case, probably you'll get more out of OpenAI's coding plan, as (from what I hear routinely) the GPT 5.6 series is thrifty with token use but as smart as the Claude 5 series: https://x.com/ArtificialAnlys/status/2085083490056589784 / https://archive.vn/3VDlN
Comment by holoduke 14 hours ago
Comment by rvba 14 hours ago
Comment by system2 14 hours ago
Comment by copperx 12 hours ago
Comment by refulgentis 15 hours ago
Comment by petercooper 15 hours ago
Comment by cedws 14 hours ago
Comment by npn 14 hours ago
Comment by andai 15 hours ago
Comment by jaggs 13 hours ago
Comment by efficax 9 hours ago
Comment by shortformblog 15 hours ago
While it won’t cover everything DeepSeek does, it handles sophisticated tasks quite well. I found it after spotting it on a chart of different models and it was listed as being near Deepseek v4 Flash’s price/performance levels.
I have noticed by the way that DeepSeek’s API has been pretty slow the past couple of days. This feels like a demand-driven move more than anything. Good thing I invested in an eGPU!
Comment by ignoramous 12 hours ago
Curious: Which one?
> MiMo-V2.5-Pro to be a pretty cost-effective alternative ... I found it after spotting it on a chart of different models and it was listed as being near DeepSeek v4 Flash's price/performance levels.
MiMo v2.5 Pro is at DeepSeek v4 Pro price level (but consumes lesser tokens per task, so cheaper overall). DeepSeek v4 Flash costs ~3x lesser than the Pro variant!
Comment by shortformblog 10 hours ago
https://tedium.co/2026/06/10/gigabyte-aorus-5060-ti-ai-box-e...
Comment by mdrzn 15 hours ago
Comment by svnt 15 hours ago
> We plan to raise the overall pricing for DeepSeek API services in the near future, with a significant increase expected. Please plan your usage accordingly. The specific pricing plan will be subject to official notice.
Comment by patates 15 hours ago
I guess this means that this powerful model for very cheap concept does not work?
Comment by mosura 15 hours ago
Their v4 flash is brilliant, and I expect they have had a huge, and expensive to handle at short notice, surge in customers as a result.
Comment by trollbridge 14 hours ago
Comment by mosura 14 hours ago
Comment by trollbridge 11 hours ago
Nearly irrelevant for US/Eastern unless you’ve got major insomnia.
Comment by f6v 14 hours ago
Remember how cheap Uber rides were? That being said, current models are absolutely incredible. And I’ll think they’re going to get cheaper when the next generation is released.
Comment by anigbrowl 14 hours ago
We plan to raise the overall pricing for DeepSeek API services in the near future, with a significant increase expected. Please plan your usage accordingly. The specific pricing plan will be subject to official notice.
Comment by storus 15 hours ago
Comment by tempoponet 14 hours ago
but compared to the $500k+ racks running the current frontier, it does give perspective.
Comment by recov 15 hours ago
Comment by vitaflo 15 hours ago
Comment by midnightbobarun 14 hours ago
Comment by GTP 15 hours ago
Comment by timpera 15 hours ago
It would be nice to have this notice on the public documentation as well.
Comment by daemonologist 4 hours ago
Comment by abdullahkhalids 15 hours ago
Comment by AtlanticThird 15 hours ago
Comment by andai 15 hours ago
Comment by LoganDark 15 hours ago
> For us, a reasonable profit means roughly this: we buy a batch of servers, and we recover the cost in about ten months. Given the risks and the upfront investment, even if we depreciate a server financially over three or five years, commercially we think a ten-month payback is enough. That is the logic behind our current API pricing. For V3.2 Flash and other models, the standard is the same: recover the cost of the equipment in ten months.
[0]: https://thechatr.ai/blog/deepseek-liang-wenfeng-investor-mee...
Comment by Readerium 11 hours ago
Comment by htrp 14 hours ago
Comment by elmer2 14 hours ago
If not, the cost outweighs whatever value I might have gotten from it.
Comment by surgical_fire 14 hours ago
I put 10 bucks on DeepSeek almost 3 months ago.
I still have 2 bucks there. I think I used so far something in the vicinity of 300M tokens total.
Even if they double their price it is still cheaper than US models on their flat rate plans, and I don't get locked out for expiring the quota.
Comment by miroljub 15 hours ago
We plan to raise the overall pricing for DeepSeek API services in the near future, with a significant increase expected. Please plan your usage accordingly. The specific pricing plan will be subject to official notice.
Now, the question remains, what does that "significant" mean? Would it still be cheaper than the competition, or did they realize they are too cheap for what they offer?
Given that many inference providers offer DS4 flash for more or less the same input and output token price as DeepSeek, they have good profits even with todays low prices.
Comment by _aavaa_ 15 hours ago
Comment by miroljub 15 hours ago
Comment by _aavaa_ 14 hours ago
Comment by faangguyindia 15 hours ago
Comment by HarHarVeryFunny 15 hours ago
Comment by jLaForest 15 hours ago
Comment by forsalebypwner 15 hours ago
Comment by apercu 15 hours ago
Please don't tell me I'm holding it wrong.
Comment by ygjb 15 hours ago
You haven't even provided enough information for us to know if you are using the right tool. Saying "the model" in relation to an LLM provider that offers several is like asking for help with using a Dewalt or Milwaukee to help assemble an cabinet. Are you cutting, drilling, hammering, screwing?
Comment by apercu 12 hours ago
They do "unstuck" me sometimes. There is value in that. But the value is inconsistent - which was my original point.
Comment by ygjb 9 hours ago
It's probably not what you want to hear, but are you sure you were holding the tool correctly? At least with this tool, you can actually ask it :D
Comment by WhereIsTheTruth 14 hours ago
- The Big Promise: Announces a major price cut for Summer 2026, powered by cost savings from a shift to Huawei chips
- The Hype Train: Launches a flash promotion cutting prices in half, follows up by announcing the 50% discount is now permanent
- The Reality Check: Summer arrives, and Huawei's chips flops
- The Retreat: Forced to quietly reorder Nvidia chips
- The Damage Control: Introduces "off peak hours" pricing to shift demand
- The Aftermath: Announces a significant price hike
Ready to IPO :^)
Comment by surgical_fire 14 hours ago
Comment by fintuner 15 hours ago
Comment by fourfire 14 hours ago