open-r1 · Issues· 340 open
Open on GitHubLocally synced open issues (discussions stay on GitHub)
- #723
New aproach to developing AGI
Updated Sep 7, 2026 - #343
in SFT script, distributed training got stuck if set `packing=false`
Updated Mar 22, 2026 - #692
RuntimeError: The size of tensor a (0) must match the size of tensor b (5120) at non-singleton dimension 1
Updated Mar 13, 2026 - #719
Dependency conflicts
Updated Feb 27, 2026 - #341
Bug in SFT script
Updated Feb 3, 2026 - #715
Is there a Chinese version of the dataset in this project?
Updated Dec 18, 2025 - #710
SFT Trains the Model on the Entire Sequence
Updated Nov 10, 2025 - #707
Forward reward always 0
Updated Oct 24, 2025 - #705
Checkpoint selection when using cot sft
Updated Oct 1, 2025 - #704
why kl = nan when grpo train?
Updated Sep 9, 2025 - #702
Request for License Information for Mixture-of-Thoughts Dataset
Updated Aug 30, 2025 - #403
Instead of rising steadily, the reward fluctuates wildly
Updated Aug 20, 2025 - #699
Does the SFT Framework Support Fine-Tuning for DeepSeek-R1-Distill-Qwen-7B
Updated Aug 18, 2025 - #698
Question about evaluating AIME24 Accuracy
Updated Aug 15, 2025 - #695
vllm prepends two BOS for LLama
Updated Aug 6, 2025