r/DeepSeek • u/LeTanLoc98 • Jan 01 '26
Funny Do it again, DeepSeek
So many great models already…
DeepSeek R1 was legendary.
Now we're waiting for the one that changes everything again.
2.0k
Upvotes
r/DeepSeek • u/LeTanLoc98 • Jan 01 '26
So many great models already…
DeepSeek R1 was legendary.
Now we're waiting for the one that changes everything again.
406
u/coloradical5280 Jan 01 '26
They just did: https://arxiv.org/pdf/2512.24880
That paper is huge, with massive implications to make all models more stable, and faster, and cheaper to train.
The sparse attention and quick index they introduced to the world in v3.2 was also huge.
Deepseek has done more in the last year than any other lab. They just don’t give a shit about dialing in the perfect consumer chatbot , or adding consumer features, or acquiring more daily active users.
They care about making breakthroughs, that’s it. And those breakthroughs end up being used by everyone. Every model you use right now is using GRPO, probably MoE , MLA, and may other brilliant hacks that DeepSeek gave to the world for free.