Programming

17978 readers

220 users here now

Welcome to the main community in programming.dev! Feel free to post anything relating to programming here!

Cross posting is strongly encouraged in the instance. If you feel your post or another person's post makes sense in another community cross post into it.

Hope you enjoy the instance!

Rules

Follow the programming.dev instance rules
Keep content related to programming in some way
If you're posting long videos try to add in some form of tldr for those who don't want to watch videos

Wormhole

Follow the wormhole through a path of communities !webdev@programming.dev

founded 2 years ago

MODERATORS

snowe@programming.dev

Ategon@programming.dev

MaungaHikoi@lemmy.nz

102

Open-R1: a fully open reproduction of DeepSeek-R1 (huggingface.co)

submitted 3 days ago by sirico to c/programming@programming.dev

9 comments fedilink hide all child comments

top 9 comments

sorted by: hot top controversial new old

[–] vrighter@discuss.tchncs.de 8 points 2 days ago

that's why gen ai models are not "open source", ever. If they were, this group would't have to "try", they could just run the build script.

Of course, the training data and software is not available. The weights are just a binary blob. It's not the source, but merely the "compiled binary"

[–] pennomi@lemmy.world 32 points 3 days ago (1 children)

This is just a proposal to make an open reproduction, nobody actually did it yet.

[–] Kuinox@lemmy.world 5 points 2 days ago

It's not only a proposal but an announcement that they are trying to do it.

[–] ericjmorey@programming.dev 17 points 2 days ago

From the article:

DeepSeek-R1 release leaves open several questions about:

Data collection: How were the reasoning-specific datasets curated?
Model training: No training code was released by DeepSeek, so it is unknown which hyperparameters work best and how they differ across different model families and scales.
Scaling laws: What are the compute and data trade-offs in training reasoning models?

These questions prompted us to launch the Open-R1 project, an initiative to systematically reconstruct DeepSeek-R1’s data and training pipeline, validate its claims, and push the boundaries of open reasoning models. By building Open-R1, we aim to provide transparency on how reinforcement learning can enhance reasoning, share reproducible insights with the open-source community, and create a foundation for future models to leverage these techniques.

In this blog post we take a look at key ingredients behind DeepSeek-R1, which parts we plan to replicate, and how to contribute to the Open-R1 project

[–] manicdave 3 points 2 days ago (1 children)

All I want is a 3gb model for the raspberry pi. 7b is too big and 1.5b is too stupid.

[–] Evotech@lemmy.world 6 points 2 days ago (1 children)

3B is probably also pretty dumb

[–] TomasEkeli@programming.dev 4 points 2 days ago (2 children)

honestly both 7b and 8b are pretty dumb as well.

[–] MadhuGururajan@programming.dev 1 points 13 hours ago

we could add so much deterministic code at 1.5GB that would start religions..

[–] Evotech@lemmy.world 1 points 2 days ago

True