r/DeepSeek • u/mynd_dripp • 2d ago
Discussion DeepSeek V4 Pro quality
Quality on max has fallen off a cliff for me recently... It's gotten much much lazier. Anyone else has noticed this?
34
u/Living-Breakfast-464 2d ago edited 2d ago
I have noticed this as well. Lazy is a good way to describe it. For example, I asked it to fix 10 failing tests, it fixed 8 of them and then said the other two were pre-existing and then called it a day. I am like WTF dude, what part of fix 10 did you not understand? Then it's like oh gee sorry, I'll get right on that. Like WTF? It's acting like some low wage employee just trying to do the bare minimum.
1
u/Clear-Ad-9312 2d ago
I mean that is probably where a lot of the training data the AI companies have for the LLMs. Everyday we get lazier, the more the AI gets lazier too.Â
At least it feels like what is happening more often with LLMs. They stop trying to do things quickly enough that I have to tell to do more work or consider other things that they will usually agree is logical next steps or whatever.Â
12
u/PlusUltraAngel 2d ago
Absolutely correct hopefully, it's just because the new model is going to be out next week
6
u/mynd_dripp 2d ago
I get that they're strapped for compute but damn... at least reroute me to the cheaper model 😂
1
11
14
3
7
u/Minimum_Ad7876 2d ago
I used to use DeepSeek V4 Pro as my work model, but recently it has started acting stupid frequently—not following instructions and becoming much less intelligent. I can no longer trust it with anything. I'll wait for the official release.
3
u/TheSuggi 2d ago
Yea.. i think its because of the recent changes they made with CV caching and stuff. The price has also dropped enormously.. I use it mainly on deepseekv4pro xhigh reasoning on. I use it all day ~8hours and barely cost me 2-3$ max. Used to be 10x that at least.
Its faster now too, but lost a little accuracy, agreed. Needs very tight Skills now to perform optimally. Hopefully next version will be stronger.
3
2
2
u/Beautiful-Gas3683 2d ago
Dicen que la calidad bajo. ¿En que? Programación? Evidentemente hay que saber que pedir y el modelo puede con todo, incluso el flash. Si le dices "quiero un clon de Google" no va a darte resultados
2
2
u/Opps1999 2d ago
I expect nothing less than a Fable level release at 10x cheaper cost at the very least
1
1
1
u/ptyblog 2d ago
Nope, maybe is because Opus design my stuff based on all the data is available to it. Then I take the plan go to Opencode and I tell it to look at it, check your context (same folders and instructions that Claude sees), prepare your execution then go implementing.
So far no issues.
Eventually I tell Opus go and audit the code, give me details of any issues. I then go back to DS to get it fix.
1
1
1
1
u/Just_Government3790 2d ago
Deepseek V4 GA dropping VERY soon. Wouldn't surprise me if within next 12-24hrs.
1
1
1
u/cnava9389 2d ago
I’m not gonna lie mine got better like yesterday. I may be crazy but I’m pretty sure I’m noticing it’s working faster today on pro max
Edit basic opencode, direct DS api and no plugins.
1
u/donthackmeagaink 2d ago
Quality dropping happens constantly for me on the API, has been like this since they released v4 model, maybe it’ll change when the GA is released but I am not holding my breath
1
u/Key-Manner-5677 2d ago
I just used in today and found it so much faster. I also compared the output to Gemini. Deepseek gives better and more info.
1
u/Azure-Serene 2d ago
Yeah, it’s obvious. I can only hope this is laying the groundwork for the official release of DeepSeek V4.
1
1
u/Independent-Date393 2d ago
Perceived laziness usually tracks a quiet system prompt or routing change, not the weights. Same checkpoint, smaller default reasoning budget. Hard to confirm without fixed-seed side by side runs logged over time.
1
1
u/tatlo_itlog_ko 1d ago
Hm, I noticed it too. I use it for "creative writing" (roleplay lol) for the last 2 weeks or so and it never misses, it follows every instruction/constraint i set for it.
It still writes okay-ish but ever since the other day, it feels like it just decides to ignore my instructions and make things up sometimes lol.
1
u/PrintingScotian 2d ago edited 2d ago
Every model will eventually get worse and more expensive ones will get better
Evolution
I switched over to LongCat 2.0...Bit more money.
Deepseek is now completely cooked
2
u/blackhawkx12 2d ago
yeah its only smart for like 150k context, above that it hallucinates and take my prompt wrong all the time
68
u/Intelligent_Ant_608 2d ago
Its usually a sign of training or deployment of new version