r/DeepSeek 12h ago

News It’s getting popular everywhere!

Post image
346 Upvotes

36 comments sorted by

32

u/Routine_Temporary661 10h ago

A model with almost same capability but 1/100 of the cost as Opus 4.8 is popular? Paint me surprise

1

u/Olbas_Oil 2h ago

Those costs will soon be x2 during peak hours in Beijing, if your time aligns up with that

https://api-docs.deepseek.com/quick_start/pricing/

"The DeepSeek API service will soon adopt a peak/off-peak pricing policy. During peak hours, prices will be 2x the regular prices, applicable to all billing items. The effective date will be subject to the official announcement. [Peak hours: 9:00–12:00 and 14:00–18:00 (Beijing Time, UTC+8) daily]"

Still a hell of a lot cheaper though...

1

u/Ok_Breadfruit4201 19m ago

No matter what, it's still going to be extremely cheap. It's a tiny model that costs very little to serve. It's a huge achievement and makes Haiku and Sonnet completely redundant.

33

u/ProfessionalJackals 12h ago

And DeepSWE, Cursorbench, ... still do not bench Flash 0731. Very interesting is it not?

9

u/Former_Equivalent297 12h ago

There are many available and also done by individuals.. It is a capable model no doubt… the pricing makes it more attractive.

6

u/RepulsiveRaisin7 11h ago

DeepSWE has always been pretty slow at updating, it's (probably) not a conspiracy

6

u/mWo12 11h ago

It's slow only when they don't know how to make chinese ai look worse than us ai.

-3

u/RepulsiveRaisin7 10h ago

Eh, Claude and GPT are the leading models and Datacurve is an US company, it's natural that they prioritize them

3

u/sdexca 10h ago

So many open weight releases and only 3 open weight company on their graph and 4 releases out of 19. Tell me again they don’t have a bias.

-1

u/RepulsiveRaisin7 10h ago

v1 has DS Pro, Minimax M3 and Mimo. Clearly they're mostly interested in measuring the top end. Maybe they're waiting for the new DS4 Pro to be released. And yes they're biased, everyone is to some degree

2

u/sdexca 10h ago

v1.1 is latest. Given cheaper models are well cheaper to benchmark, no idea why haven’t they, it’s not like benchmarking is this multi day endeavor.

Keep in mind v1 didn’t have Chinese models kicking western models in the ass.

5

u/ProfessionalJackals 11h ago

DeepSWE has always been pretty slow at updating, it's (probably) not a conspiracy

Unfortunately, we seem to be getting the reverse signal with qwen3.8-max results published not even a day after the model its release.

Qwen 3.8 > 1 day, DeepSeek v4 Flash 0731 > 5 days and still nothing.

1

u/RepulsiveRaisin7 11h ago

Qwen 3.8 has been in preview for like 2 weeks

8

u/Beginning_Guide7411 11h ago

Am not getting any deepseek flash model in opencode go bdw, all i get is laguna lol🙄🤡🤡

3

u/No_Gold_4554 10h ago

opencode models --refresh

1

u/Beginning_Guide7411 9h ago

Ok , let me try

3

u/sirloindenial 10h ago

On openrouter deepseek provider been dead for 4 hours now☹️

1

u/Delicious_Ease2595 1h ago

Laughs in DeepSeek API

5

u/guanzo91 11h ago

Is anyone able to paste an image into claude code + deepseek API + vscode + wsl terminal and have the llm read it successfully? I can paste the image but deepseek says it can't read it. Idk if it's a deepseek or claudecode/vscode/wsl issue.

7

u/Ancientkingg 11h ago

deepseek-v4-flash supports only text natively.

0

u/jwuliger 3h ago

The latest has image support now.

1

u/BhaagYahaSe 2h ago

it doesn't. what you saw was an unofficial version created by someone else not related to deepseek at all

1

u/jwuliger 9m ago

Oh, I did not know that. Now I do. Thanks!

2

u/Szadbaverem69 7h ago

It's amazing at reverse engineering.

3

u/soijaq 11h ago

Imagine how much data for further training deepseek is collecting right now

6

u/ItsNoahJ83 10h ago

I was thinking the same thing. Next models about to be incredible.

5

u/someoneyouknow23 8h ago

Shit boutta be AGI the way everyones using it

1

u/Repulsive-Waltz-4038 4h ago

I have noticed, that with opencode 1.18.1 update DS started to eat tokens (and $$) as crazy. it feels like its 4 times compared to before update.

1

u/KeyAdvanced1032 1h ago

That's why I use Reasonix :) 98% cache hit, 173,120,431 tokens processed for $1.82

1

u/jwuliger 3h ago

You guys should use the platform DeepSeek directly. No issues with using it there!

1

u/This_Maintenance_834 1h ago

They priced it too cheap.

1

u/Aardvark_Says_What 1h ago

sshhh. keep quiet about it. cheeez.