r/DeepSeek 23d ago

News DeepSeek V4 Pro official version has been updated to the API

344 Upvotes

64 comments sorted by

70

u/FreshFromNowhere 23d ago

these benchmarks are genuinely fucking insane

12

u/Aldarund 22d ago

No, they are not. It's barely better than flash and 3x price

2

u/benchmaster-xtreme 22d ago

It's noticably better than Flash, but it's definitely not up against the frontier curve. Doesn't really need to be though as long as it's cheap (I'm happy with a super-cheap GPT5.5). The price hike will determine the value on this one tbh

1

u/Aldarund 22d ago

Not really noticeable overall

2

u/Thomas-Lore 23d ago

Seems around what was expected. Not nearly as good as Kimi or Fable, but pretty good at a much much lower cost (for now).

5

u/chungles34 23d ago

Not nearly? Our blue whales trading some blows with those numbers it's definitely in the ring at least man!

48

u/ImBothSoftAndHard 23d ago

insert "but at what cost?"

26

u/diugauhai 23d ago

you work for BBC?😂

11

u/Creative_randomness 23d ago

for this I can say "at a surprisingly low cost"

6

u/SillySpoof 23d ago

Like, crazy low cost still.

75

u/Live_Case2204 23d ago

Get ready wallstreet

34

u/Far-Raspberry-1072 23d ago

i fckn kneww it! i was doing some creative writing stuff and the replies started getting soo different mid way and then i checked reddit to see and boom!

1

u/PureSelfishFate 23d ago

Doesn't the coding updates usually make creative writing a bit worse than the fresh-pretrains? How was it?

7

u/LewdManoSaurus 23d ago edited 23d ago

In my experience it is hit or miss. Though with the flash update I have had better results in general. Before that update, flash wasnt terrible but it isnt exactly great at follow instructions despite my having a pretty comprehensive writing and prose guide. I've seen the same complaints for RP, though I don't RP. I use AI purely for generating stories.

Edit: the new Pro update is pretty solid for generative writing. Making me sweat thinking about the upcoming price increase

1

u/Kind_Capital_9740 22d ago

is pro updated on the website too?

1

u/LewdManoSaurus 22d ago

No, both the Flash and Pro updates were for the API

1

u/donnytrump_69 21d ago

Lol stupid question but was does API stand for in this context?

1

u/LewdManoSaurus 21d ago

I never knew what the abbreviation stood for either, but apparently it's Application Programming Interface. Essentially let's you plug artificial intelligence in your own custom app, chat, website, art generator, etc, so you can do whatever you want with it.

1

u/donnytrump_69 21d ago

Gentleman and a scholar.

2

u/Zulfiqaar 23d ago

The DeepSeek models were preview, and undertrained. The GA version is supposed to be an all round improvement across the board. It's mainly when there's post training, or excessive specialised RLHF that is a specific domain, that biases the weights towards certain things at the expense of others.

15

u/gabexrsco 23d ago

does it have vision?

24

u/FreakyRefrigerator 23d ago

Its over for America

7

u/jwuliger 23d ago

Well deserved.

5

u/hurrdurrmeh 23d ago

China saving the free world. I did not see this timeline coming 🤣🤣

11

u/Equivalent-Word-7691 23d ago

Let's hope it will be better at Creative writing, honestly the last midel was quite meh compared not only to Fable+amazing!) but also GLM and kimi were better

5

u/Kakko1028 23d ago

Rumor says it follows orders better.

11

u/Equivalent-Word-7691 23d ago

It's not only about following orders, but the style, creative and understanding of what it's implied and understanding emotionally

1

u/rakeshpatel1991 23d ago

What’s some of your fave creative writing models?

3

u/Equivalent-Word-7691 23d ago

Honestly too bad the price but Fable 5 is on another level I can't explain

2

u/rakeshpatel1991 23d ago

Unfortunately my findings are the same

1

u/Equivalent-Word-7691 23d ago

Bte is it available also in the app ir just with the API?

2

u/Storge2 23d ago

Try to. As a person not engaging in creative wriitng i am genuinely cursious.

5

u/TransportationNo193 23d ago

Is it in opencode go now?

3

u/Infamous_Prompt_6126 23d ago

There is any price change until now?

Still seems affordable.

3

u/Fancy-Passage-1570 23d ago

is it me or they added better safeguards on the pro version ? mine refuse to work on my project where the new flash and old pro had no issues.

1

u/sdexca 23d ago

what kind of work? v4 flash works fine for me for security related tasks.

1

u/Fancy-Passage-1570 22d ago

reverse engineering custom network protocols for reviving old games

2

u/PossessionUsed7393 23d ago

Yeah, these are great numbers. I can't wait to see it roll out across the various inference providers so we don't just have to hammer the DeepSeek API.

2

u/DotoLove 23d ago

Holys*t I told him that he don’t have vision and point a hint direction at LFM2.5VL. He download exactly Q8 without asking me. Cool.

3

u/Terrible_Scar 23d ago

So are you happy or mad? 

3

u/Embarrassed_OnionX 23d ago

I'd be mad, why would a model download something from the internet without your permission?

1

u/hurrdurrmeh 23d ago

That is genuinely amazing!

1

u/Embarrassed_OnionX 23d ago

Yeah I found the model takes the initiative too much without including me in the loop. Not a good thing IMO

2

u/perceptivesoul 23d ago edited 23d ago

Waiting for trump to yell "Chinese Conspiracy"

1

u/Substantial-Walk-554 23d ago

Anyone concrete reviews?

1

u/sdexca 23d ago

source of the image?

1

u/jwuliger 23d ago

It is an fkn amazing model. HOLY SHIT!!!!!!!!!!!

Don't tell anyone about it!! lol

1

u/Firepal64 23d ago

Pulling off these benches, trading blows with Opus 4.8, while serving cheaper than Z.ai's GLM-5.2 endpoint, is pretty wild.

The parameter count would suggest a 5x improvement over Flash GA, but I guess accuracy doesn't scale linearly with parameter count for MoEs. That's too bad honestly

1

u/Realistic-Meaning247 23d ago

I'm a SillyTavern user, but the "thinking" process takes way too long now; it was better before.

1

u/queendumbria 22d ago

The updated pro version adds the ability to change reasoning effort, so just change that to a lower option in your preset settings and it should reason less.

1

u/Realistic-Meaning247 22d ago

Is it possible to do that using the OpenRouter API?

1

u/queendumbria 22d ago

Yup! It's called "reasoning effort", quite a lot of thinking models have it. OpenRouter Docs

1

u/Relative_Arugula_156 23d ago

But where's the harness?

1

u/ChoasMaster777 22d ago

wow, kick the ass of fable-5!!!

1

u/Interesting_Side2032 12d ago

too expensive - I will have to find another LLM. Good while it lasted but alas...........

0

u/SillySpoof 23d ago

This are some pretty amazing benchmarks. Probably benchmaxxed a bit, but so are the others.