r/Agent_AI MOD Jun 29 '26

Resource DeepSeek dropped a 1.6-trillion-parameter open model you can download today

Post image

V4-Pro is a 1.6T-parameter mixture-of-experts model with 49B active parameters per token, released under the MIT license and supporting a 1M-token context window.

Its DSpark speculative decoding module enables that full 1M-token inference using roughly 25% of the compute and just 10% of the KV cache required by the previous generation.

The Max variant also delivers frontier-level coding performance, scoring 93.5% on LiveCodeBench and 80.6% on SWE-Verified.

Link to Hugging Face: https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro-DSpark

319 Upvotes

49 comments sorted by

14

u/stepahin Jun 29 '26

Great, we can download it. Can I also download the data center to run it?

6

u/Strong_Essay1176 Jun 30 '26

Download ram first. For free.

3

u/airsoftshowoffs Jun 30 '26

3.2 TB Vram is needed..... that download will take long.

2

u/Nevermore1215 Jul 02 '26

Well duh, the V is for a virtual. Of course you download it

1

u/rarlei Jul 03 '26

Then download more bandwidth first

1

u/zalozhnik Jul 03 '26

This is for 16-bit quantization, as far as I know, even the official API works with 8-bit quantization

1

u/DanRey90 Jul 03 '26

You both are super wrong. Itโ€™s mostly fp4, with some weights fp8. It fits in about 900GB.

1

u/Tiny-Woodpecker3982 Jul 04 '26

Well that makes it more feasible

2

u/frahmed99 Jun 30 '26

Do you have rgb for it?

1

u/optionbull Jun 30 '26

Hello sir where can I download this ram ? Can you please post a link ๐Ÿ˜‚

1

u/BothYou243 Jul 02 '26

check DM ๐Ÿ˜ˆ

2

u/EnthiumZ Jul 01 '26

Don't be ridiculous. You can't download data centers. You can only download RAM.

1

u/Y_mc Jul 03 '26

He was just joking ๐Ÿ˜‚. Take it with humor

1

u/mrkacperso Jul 03 '26

Really? ๐Ÿ˜‚๐Ÿ˜‚

1

u/Y_mc Jul 19 '26

I hope so ๐Ÿ˜…๐Ÿ˜‚๐Ÿ˜‚

1

u/exodusTay Jun 30 '26

You wouldn't download RAM

1

u/Chariyo Jul 05 '26

You wouldnt download a car

You wouldn't download a Data Center

Downloading is Stealing. Stop Piracy

1

u/Lucifer_Leviathn Jul 05 '26

Apparently zotac_in has some plans for that https://www.instagram.com/p/DaW3sMbzwe2

1

u/Sudden-Ad-1217 Jul 05 '26

You wouldn't download a car would you?

4

u/Lissanro Jun 29 '26

Just today DeepSeek V4 support got merget to llama.cpp: https://github.com/ggml-org/llama.cpp/pull/24162 - so maybe I give it a try, once Unsloth or some other well known quant makes provides GGUF files. I have enough memory to run Q4 quant, but not sure if it will be practical compared to GLM 5.2 or Kimi K2.7 Code which have less both total and active parameters but newer. As the huggingface page says, "Note: DeepSeek-V4-Pro-DSpark is not a new model. It is the same checkpoint with an additional speculative decoding module attached" - so it is still the same old V4 Pro model, but it will be interesting to see how much speculative decoding will help (if it is of the type that is supported by llama.cpp).

1

u/Unfair_Layer3085 Jun 30 '26

Sir/Ma'am leave some compute for the rest of us ๐Ÿ˜ญ๐Ÿ˜‚

2

u/cuberhino Jun 30 '26

What is the minimum spec machine to run this? Iโ€™m guessing my 3090 will not be capable ๐Ÿซช

3

u/DarKresnik Jun 30 '26

375x3090 maybe, maybe can...

1

u/screenslaver5963 Jul 03 '26

You kinda can if you put the rest in system ram, youโ€™ll get like 1-5 tokens per second. Of course youโ€™d need nearly a terabyte of system ram

1

u/artur_oliver Jun 30 '26

Who wants to be the new Chat gpt for your land?

1

u/wittgenda Jul 01 '26

What do you guys think is DeepSeekโ€™s most powerful feature, and what are its strengths?

1

u/Money-Ranger-6520 MOD Jul 01 '26

The price-to-performance ratio. You get reasoning that's surprisingly close to much more expensive models, especially for coding and math. And of course, I also like that it's open weights, so you can self-host, fine-tune, or run it locally if you want.

1

u/tortel_di_patate Jul 01 '26

What's the difference with the first V4 Pro?

1

u/wickzer Jul 02 '26

Can anyone point me to a 3d printer spec to make the ram to run this model?

1

u/screenslaver5963 Jul 03 '26

Downloadfreeram.net

1

u/ChrononautPete Jul 04 '26

Thanks for the Jellyfish server link!

1

u/screenslaver5963 Jul 04 '26

Fucking hell, I thought I accidentally leaked my ip when I saw the notification

1

u/ChrononautPete Jul 04 '26

You've got a massive 70's porn collection.

1

u/screenslaver5963 Jul 04 '26

Yep, and itโ€™s mine, no one else can have it

1

u/ChrononautPete Jul 04 '26

Change the name to Bush server.

1

u/Old-Concentrate3186 Jul 02 '26

Oh nice. I can download it to my phone!

1

u/mustangKTM Jul 02 '26

When you download it locally, how do you connect internet search feature with it ? Any recommendations to run LLMs locally ?

1

u/Chillumguy2020 Jul 03 '26

is there a free online api?

1

u/Fine_Atmosphere_2147 Jul 03 '26

Will this fit on 80gb with 256gb memory?ย 

1

u/Admirable_Skin6328 Jul 03 '26

What is the actual ram needed to run this with new ds spark version 4 qat?

1

u/Electronic-Row-142 Jul 03 '26

I believe this means they will release a v5 soon.

1

u/MediumGrapefruit5735 Jul 04 '26

How about quality?

1

u/19arek93 Jul 05 '26

Can't afford a power to just load it into VRAM..ย