pytorch-lightning
https://github.com/pytorchlightning/pytorch-lightning
Python
The lightweight PyTorch wrapper for high-performance AI research. Scale your models, not the boilerplate.
Triage Issues!
When you volunteer to triage issues, you'll receive an email each day with a link to an open issue that needs help in this project. You'll also receive instructions on how to triage issues.
Triage Docs!
Receive a documented method or class from your favorite GitHub repos in your inbox every day. If you're really pro, receive undocumented methods or classes and supercharge your commit history.
Python not yet supported8 Subscribers
Add a CodeTriage badge to pytorch-lightning
Help out
- Issues
- Loading large models with fabric, FSDP and empty_init=True does not work
- AWS Trainium fails number of device validation when using more than 1 accelerator on the instances
- OnExceptionCheckpoint: training resumes if ckpt found, even if no ckpt_path provided
- Checkpoint every_n_steps reruns epoch on restore
- Multi-node Training with DDP stuck at "Initialize distributed..." on SLURM cluster
- Differentiate testing multiple sets/models when logging
- Construct objects from yaml by classmethod
- Current FSDPPrecision does not support custom scaler for 16-mixed precision
- Please make it simple!
- Multi-gpu training is much lower than single gpu (due to additional processes?)
- Docs
- Python not yet supported