litgpt · Issues· 290 open
Open on GitHubLocally synced open issues (discussions stay on GitHub)
- #2223
No checkpoints saved during pretraining in Google Colab despite log.log_model and save_interval set.
bugUpdated Sep 8, 2026 - #2323
LLM.generate() can silently select a token excluded by top_k on CPU when sampling probabilities are float16
Updated Sep 8, 2026 - #1242
False positive warning about mixed precision in `merge_lora.py`
Updated Sep 7, 2026 - #2306
Reang
questionUpdated Sep 4, 2026 - #2319
litgpt pretrain crashes or deadlocks when a rank's val_dataloader yields zero batches (multi-GPU FSDP)
Updated Sep 4, 2026 - #1102
Make `save_hyperparameters()` robust against different CLI entry points
bughelp wantedUpdated Aug 30, 2026 - #2307
Speculative decoding drops the first generated token
Updated Aug 25, 2026 - #2302
Stop sequence matching mishandles overlapping prefixes and buffered tokens
Updated Aug 19, 2026 - #2189
Unsafe Checkpoint Loading - This Is Bad
Updated Jul 21, 2026 - #2158
Support for SSM models (Mamba, Mamba2)
enhancementUpdated Jul 20, 2026 - #2116
Initial and final evaluation in `finetune` scripts do not accumulate over devices
bugUpdated Jun 13, 2026 - #1871
failure converting pretrained litgpt checkpoints to HF format: a reproducible example
bughelp wantedUpdated Jun 13, 2026 - #2191
Gradient Clipping Doesn't Work in Finetuning
questionUpdated Jun 13, 2026 - #1221
Calculate loss at beginning and end of training
enhancementUpdated May 23, 2026 - #1082
Add TinyStories to the pretraining docs
documentationUpdated May 23, 2026