Issues du dépôt
Lightning-AI/pytorch-lightning
Pretrain, finetune and deploy AI models on multiple GPUs, TPUs with zero code changes.
Issues
Ouverte
LayerNorm / BatchNorm fp16 behavior is different in Pytorch Native and Deepspeed
bughelp wantedstrategy: deepspeed
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
3 commentaires1 réaction0 personne assignée
Ouverte
BUG when Trainer.test() with deepspeed stage 3
bughelp wantedrepro neededstrategy: deepspeed
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
2 commentaires2 réactions0 personne assignée
Ouverte
Schedule in PyTorchProfiler doesn't work
bughelp wantedprofiler
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
6 commentaires1 réaction0 personne assignée
Ouverte
HPC Resubmit resume on most recent epoch checkpoint
checkpointingenvironment: slurmfeaturehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
11 commentaires2 réactions0 personne assignée
Ouverte
Add `enable_device_summary` flag to disable device printout
callbackfeaturegood first issuetrainer: argument
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
15 commentaires3 réactions1 personne assignée
Ouverte
Multiple GPU per node could fail silently with KubeflowEnvironment
bugenvironment: kubeflowhelp wanted
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
22 commentaires2 réactions1 personne assignée
Ouverte
Race condition with SLURM restart
environment: slurmfeaturehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
4 commentaires0 réaction0 personne assignée
Ouverte
Use :emphasize-lines: in sphinx docs to highlight code.
docsgood first issuepriority: 1
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
11 commentaires2 réactions0 personne assignée
Ouverte
Support batch size scaling with dataloaders passed directly to `fit()`
data handlingfeaturehelp wantedtuner
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
3 commentaires0 réaction0 personne assignée
Ouverte
Use `FutureWarning` instead of `DeprecationWarning` for deprecation warning
featuregood first issue
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
6 commentaires2 réactions1 personne assignée
Ouverte
Tagging discussion with release number
docsfeaturehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
4 commentaires0 réaction0 personne assignée
Ouverte
[RFC] Default to infinite epochs, not 1000
help wantedlet's do it!trainer: argument
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
18 commentaires0 réaction0 personne assignée
Ouverte
Distributed mode overwrites the user's choice for dataloaders shuffling
data handlingfeaturehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
8 commentaires0 réaction0 personne assignée
Ouverte
ModelCheckpoint: save_top_k > 1 cannot recognize ordering of models from ckpt names
checkpointingfeaturehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
18 commentaires1 réaction0 personne assignée
Ouverte
Allow shuffling when overfit_batches is active
featurehelp wantedrefactor
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
19 commentaires0 réaction0 personne assignée
Ouverte
Progress bar doesn't show up on Kaggle TPU with `num_workers` greater than `0`.
accelerator: tpubughelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
33 commentaires0 réaction0 personne assignée
Ouverte
Syncing the log_dir across ranks is not valid with multiple nodes
bugdistributedhelp wantedloggingpriority: 1
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
3 commentaires0 réaction1 personne assignée
Ouverte
AdvancedProfiler: ValueError: Attempting to stop recording an action (run_test_evaluation) which was never started.
buggood first issuehelp wantedpriority: 1
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
18 commentaires0 réaction0 personne assignée
Ouverte
Save training metadata with the fault tolerance checkpoint
fault tolerancehelp wantedlet's do it!
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
3 commentaires0 réaction0 personne assignée
Ouverte
[RFC] Profiler Metrics
featurehelp wantedlet's do it!profiler
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
3 commentaires0 réaction2 personnes assignées