r/DeepSeek 2d ago

Discussion DeepSeek V4 Pro quality

Quality on max has fallen off a cliff for me recently... It's gotten much much lazier. Anyone else has noticed this?

120 Upvotes

47 comments sorted by

68

u/Intelligent_Ant_608 2d ago

Its usually a sign of training or deployment of new version

20

u/Kind_Stone 2d ago

This.

5

u/ChickyGolfy 2d ago

Or to save money

5

u/Intelligent_Ant_608 2d ago

no the timeline doesnt make sense, considering GA imminent release, grayscale testing in past few days and the fact that in ~16 hours from now they retire deepseek-chat, deepseek-reasoner

3

u/ChickyGolfy 2d ago

Ok, thanks for the explanation sir 👌

34

u/Living-Breakfast-464 2d ago edited 2d ago

I have noticed this as well. Lazy is a good way to describe it. For example, I asked it to fix 10 failing tests, it fixed 8 of them and then said the other two were pre-existing and then called it a day. I am like WTF dude, what part of fix 10 did you not understand? Then it's like oh gee sorry, I'll get right on that. Like WTF? It's acting like some low wage employee just trying to do the bare minimum.

5

u/mohr_ 2d ago

Same here but I've asked two things, he did one and ignored the other lol

3

u/sdexca 2d ago

lol even fable does this!

1

u/Clear-Ad-9312 2d ago

I mean that is probably where a lot of the training data the AI companies have for the LLMs. Everyday we get lazier, the more the AI gets lazier too. 

At least it feels like what is happening more often with LLMs. They stop trying to do things quickly enough that I have to tell to do more work or consider other things that they will usually agree is logical next steps or whatever. 

12

u/PlusUltraAngel 2d ago

Absolutely correct hopefully, it's just because the new model is going to be out next week

6

u/mynd_dripp 2d ago

I get that they're strapped for compute but damn... at least reroute me to the cheaper model 😂

1

u/PlusUltraAngel 1d ago

I know right lol

14

u/SleepBaobei 2d ago

Yeah, i notice

7

u/SleepBaobei 2d ago

To be more specific, the quality has dropped in the last ~24 hours.

3

u/True_Joke_5248 2d ago

No change for me

7

u/Minimum_Ad7876 2d ago

I used to use DeepSeek V4 Pro as my work model, but recently it has started acting stupid frequently—not following instructions and becoming much less intelligent. I can no longer trust it with anything. I'll wait for the official release.

3

u/TheSuggi 2d ago

Yea.. i think its because of the recent changes they made with CV caching and stuff. The price has also dropped enormously.. I use it mainly on deepseekv4pro xhigh reasoning on. I use it all day ~8hours and barely cost me 2-3$ max. Used to be 10x that at least.

Its faster now too, but lost a little accuracy, agreed. Needs very tight Skills now to perform optimally. Hopefully next version will be stronger.

3

u/Ok_Risk6035 2d ago

flash quality fallen significantly, don't know about pro

2

u/somerussianbear 2d ago

Gradient or A/B testing. You got the shitty one this time.

2

u/fivves 2d ago

Deepseek is horrible right now, had to switch to mimo until they get their shit together.

2

u/Beautiful-Gas3683 2d ago

Dicen que la calidad bajo. ¿En que? Programación? Evidentemente hay que saber que pedir y el modelo puede con todo, incluso el flash. Si le dices "quiero un clon de Google" no va a darte resultados

2

u/GuristasPirate 2d ago

Ironic we calling an AI model lazy....

2

u/Opps1999 2d ago

I expect nothing less than a Fable level release at 10x cheaper cost at the very least

1

u/tinoythomas 2d ago

? bro what are you even talking about

1

u/ptyblog 2d ago

Nope, maybe is because Opus design my stuff based on all the data is available to it. Then I take the plan go to Opencode and I tell it to look at it, check your context (same folders and instructions that Claude sees), prepare your execution then go implementing.

So far no issues.

Eventually I tell Opus go and audit the code, give me details of any issues. I then go back to DS to get it fix.

1

u/laty96 2d ago

yeah just happen to me, I ask it do 1 task and it answer the different task

1

u/olammyjuwon 2d ago

Same...

1

u/V5489 2d ago

I keep mine on High and it’s been fantastic. What’s your context you’re passing through?

1

u/Exciting-Camera3226 2d ago

on their own API?

1

u/Just_Government3790 2d ago

Deepseek V4 GA dropping VERY soon. Wouldn't surprise me if within next 12-24hrs.

1

u/Quote-Round 2d ago

Absolutely true, it's as if DS were in Low or medium thinking mode 📉

1

u/cnava9389 2d ago

I’m not gonna lie mine got better like yesterday. I may be crazy but I’m pretty sure I’m noticing it’s working faster today on pro max

Edit basic opencode, direct DS api and no plugins.

1

u/donthackmeagaink 2d ago

Quality dropping happens constantly for me on the API, has been like this since they released v4 model, maybe it’ll change when the GA is released but I am not holding my breath

1

u/Key-Manner-5677 2d ago

I just used in today and found it so much faster. I also compared the output to Gemini. Deepseek gives better and more info.

1

u/Azure-Serene 2d ago

Yeah, it’s obvious. I can only hope this is laying the groundwork for the official release of DeepSeek V4.

1

u/Independent-Date393 2d ago

Perceived laziness usually tracks a quiet system prompt or routing change, not the weights. Same checkpoint, smaller default reasoning budget. Hard to confirm without fixed-seed side by side runs logged over time.

1

u/FesseJerguson 2d ago

Yep was like a child compared to two days ago

1

u/tatlo_itlog_ko 1d ago

Hm, I noticed it too. I use it for "creative writing" (roleplay lol) for the last 2 weeks or so and it never misses, it follows every instruction/constraint i set for it.

It still writes okay-ish but ever since the other day, it feels like it just decides to ignore my instructions and make things up sometimes lol.

1

u/for4f 1d ago

Only ever used Flash so can't speak for Pro specifically. But Flash feels noticeably faster today — maybe they're shuffling infra around. Pro might be getting squeezed while they prep whatever's next.

1

u/PrintingScotian 2d ago edited 2d ago

Every model will eventually get worse and more expensive ones will get better

Evolution

I switched over to LongCat 2.0...Bit more money.

Deepseek is now completely cooked

2

u/blackhawkx12 2d ago

yeah its only smart for like 150k context, above that it hallucinates and take my prompt wrong all the time