huggingface / huggingface/open-r1

How many resources are required to train deepseek r1 671b using grpo?

Open
#413 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
26.5k
Forks
2.5k
PR merge metrics
No merged PRs in 30d

Description

.

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue names DeepSeek R1 671B and GRPO but provides no files, tests, or entry points. Start by clarifying the hardware, precision, sequence length, and training setup needed for the estimate. Done means documenting a supported, evidence-based resource estimate for this training scenario.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning
Issue type
Documentation
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
15/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.