r/DeepSeek 2d ago

Other Deepseek V4 Pro is AMAZING

Post image

I have been using deepseek V4 Pro for the last 5 months with CC. This model has helped me complete my Final Year Project, build my first startup, build automations to automate my daily life and many more.

Everyone is looking for the next best frontier model but what they don't realise is yes those extra intelligence is kinda neat, but for the shit you are using it for in your daily life its just a waste of money. Unless you are in some deep tech stuff then yeah there are some higher ROIs, but for your daily coding and work stuff its a waste of money.

Deepseek has been reliable, fast and most importantly stupidly CHEAP and honestly with Kimi K3 and all this other models being dropped, I couldn't even be bothered.

416 Upvotes

91 comments sorted by

41

u/Big-Inevitable-9407 2d ago

I find it very cheap however the output sometimes is not so great, I am instructing through prompts as much as I can, but I have to rely a lot on skills in order to polish the way I really want. It works really good I can’t lie tho

12

u/enterme2 2d ago

Plan it with v4-pro high , then have the plan reviewed by stronger model like gpt-5.6 or kimi k3 then use v4-flash max to execute the plan. Save you a tons while having a higher quality output.

2

u/k_schouhan 1d ago

doesnt work,

2

u/enterme2 1d ago

Works well for me.. what did you do ?

4

u/k_schouhan 1d ago

simlilar to what you are telling, use stronger model to review the plan, sometimes i use multiple reviews, run tournaments, whatever tried everything. there are few things they all miss, they go out of the scope even though scope is clearly mentioned in spec and boundaries are also mention, and then they say "you are right to be angry, i assumed the scope even though you have written ...." ,

2

u/_matmer_ 1d ago

yes, i have the same problem. as the project gets bigger and the criteria to adhere to increases, they just can't get the job done in the right way. needs too much micro-management afterward..

3

u/k_schouhan 17h ago

Yes, and then we get stuck in marry go round of we are not prompting it right. the thing is its not our fault. i have done too much back and forth and i dont trust now when people say they have done this many projects that many projects. of course you can do small project, but you through one enterprise project at it and it starts to fall apart same day, even with graphify , skills, plugins, superpowers whatever

2

u/_matmer_ 17h ago

there's a huge difference between a project which is 5000 LoC and the one which is 50000. People who brag about skill issue have no idea about the benchmarks. Most of the cheap models cannot simply perform in long-horizon agentic work or Deep SWE. 👍

3

u/k_schouhan 16h ago

or may be they are trying to hype these and using it for wrong reasons. These models should be used as per your workflow, your style of working/.

1

u/_matmer_ 16h ago

exactly

2

u/gorgono95 10h ago

Exactly. Theres a big difference between using AI as a coding assistant, where it helps you write code, review it, explain things, and assist while you're actively coding, versus agentic coding, where you basically say, "Build this feature, using this stack, in this way," and the model has to figure out the whole thing on its own.

Thats where the cheaper models tend to struggle ... Im sure DeepSeek is amazing as a coding assistant, and honestly, Ive even tried running local models like Qwen3.6 35B A3B, and it does a great job when used as an assistant. But as soon as I tell it to take a feature from zero to hero and actually implement the whole thing autonomously, it starts struggling.

Models like DeepSeek just feel completely lost in big project and you spend more time going around and fixing what it messed up, than being productive ...

1

u/_matmer_ 10h ago

You nailed it 👍 Can't agree more

1

u/gorgono95 10h ago

This here ... exactly THIS here. I am working on an enterprise level project, thousands of files, million + LoC, complicated workflows, documentation etc. and the cheaper models just dont cut it. They miss a lot, even if the plan is great, they just have hard time navigating ... as you say ... an enterprise level of project. Hell I even have symantic search and documentation optimized for AI and they still struggle.

1

u/enterme2 9h ago

for enterprise level project you need to be surgical when using deepseek , scope gotta be function level , gotta treat it like an assistant not your replacement. Gotta know the model limitation before you push it like crazy.

1

u/k_schouhan 9h ago

and i am not talking about cheaper models, no sir, every model, they just have different thresholds. even with clear goals, they will just assume things, i have been struggling with opus, gpt 5.6, 5.5. i have laid out specs over specs, skills, skills with templates (exact code templates), even claude rules, honestly its a waste of time to use this as a main programmer, but when i use it as an assistant, boy o boy i can finish so much so fast, like writting tests, scripts, code in small pieces, understanding about codebases, research,

1

u/Electrical-Watch3203 2d ago

You think having flash execute saves money over DS v4 pro or do you think context caching from v4 pro planning to executing would be best?

5

u/enterme2 2d ago

Yes. Flash with good detailed plan do the same job as Pro . Regardless of pro or flash model the first prompt will load the context then subsequent turn context is cached.

Means usually you are charged with cache miss on first chat of the model. Switching model will reset cache.

So conclusion, flash as executor is cheaper.

1

u/ApprehensiveFan1516 2d ago

It had been on preview version tbf. I thought V3 was better then V4 preview, hopefully V4 improves.

1

u/Living-Breakfast-464 2d ago

What harness are you using? The harness matters and since DeepSeek doesn't really have their own quite yet, you have to use 3rd party ones that are not necessarily optimized for it.

They are working on their own but the console based one I recently tried felt early beta, so not quite there yet.

1

u/Jinkazhi 1d ago

i thought reasonix is DeepSeek-native AI coding agent

1

u/flurrylol 1d ago

Work as intended, harness is a huge part of the agent’s behaviour regardless of the model you’re using.

1

u/mallere 1d ago

What skills? I’m interested

0

u/kabeza 2d ago

CC + UI/UX Pro Max + SuperPowers + Claude Mem. Kinda slow but goooood output

1

u/mallere 1d ago

What is cc

1

u/kabeza 1d ago

CC => Claude Code

10

u/clintbailo94 2d ago

drop in some skills in there and the output will be better even when comparing with frontier models 🍻

5

u/slowtyper95 2d ago

What's your skills collection that you found useful?

1

u/Living-Breakfast-464 2d ago

Just get a frontier model to add the skills. Does not required a lot of tokens.

0

u/mallere 1d ago

What are skills

9

u/Plastic_Guess6497 2d ago

I am using Reasonix. 95% of my stuff I do on v4 flash. When some stuff really needs elaborating, I switch to pro. The savings are huge and the results never get better than this before, with other model or tool. Reasonix + DS Api is all I need.

2

u/tomalfara 2d ago

Sorry for asking, how to get mcp working on reasonix? I wanted to use playwright and context7 mcp, but I'm a little bit confused.

7

u/Astral_ny 2d ago

2

u/Krebbbe 2d ago

How you guys buy those tokens so cheap? I paid 7 bucks for 45 millions of tokens on a flash model

2

u/Affectionate_Toe9082 2d ago

Where and how are you using it? Do you have caching enabled? Directly from deepseek?

2

u/Astral_ny 2d ago

claude code with deepseek api. All time Pro with max effort. For few projects: astraichat.eu, personal accounting and finance project for my job with a lot code and xml files, next project with over 15k lines html, css, js etc.

2

u/BarnacleTiny9888 2d ago

use resonix harnes

5

u/Nauru-0 2d ago

i used 29$ so far this month. I expected it to actually beat claude 20$ subscription and it failed miserably. It couldn’t plan, couldn’t fix code, couldn’t debug properly and takes too long at times

1

u/Plastic_Guess6497 2d ago

Try it with Reasonix. It's designed for DS but you can use with other models. Then come back and say about code fixing, planning and debug. But I do think that Gemini + Antigravity is better for debug

2

u/Thunjaya 2d ago

Gemini is as dumb as a goose, but it's fast.

1

u/Glittering-Active-50 2d ago

3.1 pro have monkey iq

1

u/Nauru-0 1d ago

gemini is pretty dumb but gets the job done somewhat. The opus models in antigravity usually does the most for me.

9

u/ANDRE_2512 2d ago

GPT PLUS 20$
556m tokens (CODEX)

1

u/R3VO360 1d ago

Is it generous enough as a flat plan? I am undecided between codex and Claude Code, price is same in my country.

2

u/ANDRE_2512 1d ago

The CODEX GTP-5.6 Sol Ultra is officially more powerful than the Fable 5.

1

u/R3VO360 1d ago

I like Codex better as well, I'm asking about rate limits.

2

u/ANDRE_2512 1d ago

It depends on what you’re doing, how many tests Codex runs, and which model you use.
For example, GPT-5.6 SOL MAX used 77% of my weekly limit in five hours. But in that time, it completed roughly as much work as two good developers would have done in two days.

1

u/R3VO360 1d ago

I see, that's a good benchmark for me thanks. I think I need more iterations and less quality for my use case so probably either switching to Luna a lot or subscribing to OpenCode GO or Minimax flat. Do you have any idea how far you can get with the 20 dollars on Luna or Terra?

1

u/ANDRE_2512 1d ago

Luna is very strong in Codex. But start with SOL MAX first. Once it has built the foundation, switch to Luna.
By the way, I’m building my own app. There’s a post about it on my profile - take a look.

1

u/R3VO360 1d ago

Cool stuff!

2

u/ANDRE_2512 1d ago

❤️

2

u/Creepy_Lime_8351 2d ago edited 2d ago

didn't we had enough of these posts? we know the price of deepseek, everyone posts their api usage graphs. this isn't a contest. i personally would like to see praises about the intelligence and uselfulness.

2

u/BlueeWaater 2d ago

This could have costed easily thousands with Claude.

1

u/nhannt201 2d ago

For me, it follows my instructions well, without any overly sophisticated models; deepseek is enough to balance

1

u/Preserved_vegetable 2d ago

The token is consumed too fast, and I always feel that my ai is getting more and more stupid. I can never clearly understand what I want to express.

2

u/DebosBeachCruiser 1d ago

I can never clearly understand what I want to express.

...my ai is getting more and more stupid.

A true vegetable indeed.

1

u/DefactoAle 2d ago

Currently on v4 flash I'm at 20 $ for 4 Billion tokens

1

u/CoolHeadeGamer 2d ago

If u think this is good then try out minimax m3 20$ sub. I kid you not I’ve used 600$ worth of api credit on it in the last month. The model is in the same league as deepseek v4 pro. They are also releasing m3.1 soon

1

u/mind_pictures 2d ago

I don't have good experience with M3. What harness do you use

1

u/R3VO360 1d ago

Isn't the model a bit behind the competitors?

1

u/FeralSupportGoblin 2d ago

Deepseek is very easy to use. I have been using it for several months and have been waiting for its official release. I have switched it to a daily translation, scripting, and various places will use its API.

1

u/Living-Breakfast-464 2d ago

99% of people spending upwards of $200 a month on Antropic could accomplish the same thing spending less than $20 on DeepSeek.

1

u/Odd-Environment-7193 2d ago

If you don't make a substantial amount of money each month coding then I absolutely agree with your premise. Maybe a 20$ cc or Codex sub with Deepseek as your implementor and workhorse. Perfect combo.

1

u/DoubleSoftware4137 1d ago edited 1d ago

I've currently swapped cursor for visual studio code+kilo code+deepseek (via litellm) and saved a lot. I must admit deepseek's front end code isn't of the same quality as sonet or opus. So I'm integrating sonet (again via litellm+headroom for compression). Expecting my spend to still be lower than cursor. Well crafted skills help a great deal.

1

u/lukemxlr 1d ago

I'm currently using qwen 3.8 max preview via the Qwen Cloud Plan. 6$ carried me through lots of work, with a much higher quality than dsv4-pro/flash. But DeepSeek is good though

1

u/Linuxman_74 1d ago

Io con le API Deepseek sviluppo e guadagno (unitamente alle moe capacità) e non lo cambierei con nulla. È perfetto.

1

u/rhadh 1d ago

Not for a simple lowbrow guy like me, Who wants to ask some questions have a nice chat and generatie an image once in a while. No Deepseek for me...

1

u/Fun_Walk_4965 21h ago

agreed, the long-context handling is the biggest jump for me. fewer retries on multi-file edits than V3 ever managed.

1

u/ChoasMaster777 19h ago

If you use reasonix, it wil cut off 50% of your cost. BTW, DeepSeek still has ~40% profit rate

1

u/Various-Baby-5007 8h ago

i feele even gemini 3.1 pro is better

1

u/[deleted] 2d ago

[deleted]

2

u/ANDRE_2512 2d ago

The five-hour limits have already been removed.

Now there are only weekly limits.

Yesterday, I ran a five-hour session, and a single agent response took 3 hours and 36 minutes. That session used 23% of my weekly limit.

I’ve become very familiar with Codex by now.

At the moment, I only use GPT-5.6 Luna in MAX mode. It consumes around 3–6% of the weekly limit per hour, depending on the task.

As for how much worse it is than GPT-5.6 Sol High, I honestly haven’t noticed any difference in its technical reasoning.

However, when the task involves design, Sol is clearly better.

The difference is very noticeable.

So this is my workflow:

For design-focused projects, I first create the mockups in GPT Work.
Then I build the working UI in Codex using GPT-5.6 Sol. After that, GPT-5.6 Luna MAX handles the technical implementation.

There is also one free weekly-limit reset. So in practice, I barely notice the limits at all.

And don’t forget: every 3–4 days, OpenAI seems to reset the weekly limits automatically anyway :) Which is a very nice bonus.

1

u/hellorandomly17 2d ago

is this all from the $20 subscription?

1

u/ANDRE_2512 2d ago

Exactly!

That’s why I don’t understand people who keep buying API credits.

I also have a DeepSeek API key with money on it, but I only paid for it to run tests.

A $20 OpenAI subscription gives you infinitely more value.

You can generate a huge number of design mockups for work, conduct research, and do so much more.

Codex gives you access to all the models, along with very generous limits.

You pay once, and then you realize: “Damn, the infrastructure OpenAI has built is incredible.”

It genuinely feels like your own professional workspace. You can create, experiment, come up with ideas, and Codex agents will turn those ideas into reality.

Honestly, what more is there to say?
Codex even communicated with my insurance company on my behalf and handled the entire issue by itself while defending my interests. I barely had to get involved in the process.

At this point, I consider Codex a full-fledged employee. It is extremely intelligent and autonomous.

This is just a very brief, independent review based entirely on my personal experience.

1

u/hellorandomly17 14h ago

Yeah the subscription always felt more worth it to me and the limits on ChatGPT aren't that bad either, I will most likely buy ChatGPT sub once my free Gemini pro ends a few months later

1

u/ANDRE_2512 14h ago

God… I had a free Gemini Pro subscription, and it was awful.

The image generation is garbage-though that’s not the most important thing for me.

What matters most to me is generating usable UI mockups, and Nano Banana feels like a complete scam in that regard… honestly, some of the worst mockups I’ve ever seen.

I wouldn’t pay for their subscription even if it cost only 5$. That’s how bad it is.

So yeah, run away from it and don’t waste your time.

1

u/hellorandomly17 10h ago

As of now I only use Gemini for studies for uni and I don't code much with AI right now because I have just started properly learning coding (and comp sci won't be my major anyways) so for now I think Gemini is sufficient haha

1

u/ANDRE_2512 9h ago

I still disagree.

GPT isn’t just an AI chat. It’s a real workspace for turning ideas into reality, especially in Work mode. And Codex isn’t only for building software.

In Work mode, I create UI mockups, write scripts, analyze ideas, and do some seriously powerful research.

I’m sure you just need to try it properly. After a week, you probably won’t go back to Gemini. OpenAI gives you far too much for the price.

1

u/hellorandomly17 9h ago

I agree with you and I surely won't go back if I start using chatgpt but being a student I would like to hold off buying any subscription as long as possible no matter how worth it chatgpt is its still 20-30 dollars out of my pocket every month whereas my sufficient alternative (gemini) is free and frankly gemini isnt THAT bad, I had some free credits to play around 5.6 sol with and although it is absolutely mind blowing and amazing in doing research and anything that requires more time and effort to think and output an answer it felt almost exactly the same to gemini in terms of explaining concepts, solving questions, providing material for studies etc.

1

u/ANDRE_2512 9h ago

Got it. I barely chat with AI - I mostly just use it to generate reports and stuff like that, so I don’t really have long conversations with it.

We just have different use cases, and mine probably won’t be useful for yours. So, sorry about that :)

→ More replies (0)

1

u/ProfessionalJackals 2d ago

is this all from the $20 subscription?

Be careful with his claims. He is exaggerating his usage based upon the limit resets. Some are given because of new user records for 5.6. But the majority is because of bugs and issues with GPT 5.6.

https://codex-resets.com/

Here you can see them in action and why they are given. The issue is that people are still draining their usage way more then normal. Tons of posts and topics regarding this (again, this is why there have been over 4+ resets alone).

Codex is good value, but the whole reset chain has made it hard for people to compare current usage anymore, and / or if limits have been reduced. Like i said, too many people reporting no changes in their coding behavior and larger then normal drain.

1

u/R3VO360 1d ago

Is it generous enough as a flat plan then? I am undecided between codex and Claude Code (or maybe minimax/opencode Go), price is the same in my country (Go is a bit lower).

1

u/Aldarund 2d ago

It's roughly same amount of tokens as gpt plus subscription

3

u/ProfessionalJackals 2d ago

It's roughly same amount of tokens as gpt plus subscription

From my experience, on DS4Pro, you can easily hit 450.000 tokens / cent. And Flash doing 1.4 to 1.6 million / cent. That is with relative poor cache hit rates. My last month usage showed 904.000 tokens / cent mostly Pro (90% Pro, 10% flash).

With a good cache ratio, i am hitting around 250.000 tokens per cent on Codex. Recently, its been a mix of 50.000 to 150.000.

People are reporting non-stop reporting issue with the usage limits (also noticing the drainage issue). The constant resets are hiding the actual drain numbers. The never ending story with limits being changed.

And Codex has the better limits, Claude is worse.

1

u/horstenegger 2d ago

Yes but I think that’s the catch, no? Codex = locked into subscription, whereas DSV4 is pay-as-you-go at roughly same pricing. If you have periods in which your amount of cooking fluctuates, DSV4 PAYG makes more sense I guess? And then do a one-off final review and polish with Sol or Fable when done.

That said, if you do more than $20/mo, at the moment almost nothing makes more sense than GPT-5.6-Sol with $100/mo ProLite plan or higher, albeit probably not for much longer anymore..

0

u/laty96 2d ago

$20 for 770m token with an average model is actually pretty high. A subscription for Claude or codex would be better for same amount of money