Issues du dépôt
Lightning-AI/pytorch-lightning
Pretrain, finetune and deploy AI models on multiple GPUs, TPUs with zero code changes.
Issues
Ouverte
Runtime events with support for custom handlers
featurehelp wantedlet's do it!logging
Label adapté aux débutantsGuide de contribution disponible
9 commentaires4 réactions1 personne assignée
Ouverte
Move reload_dataloaders_every_n_epochs to the DataHooks class
data handlingdeprecationdesignfeaturehelp wantedlet's do it!
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
5 commentaires4 réactions1 personne assignée
Ouverte
Pruning callback causes GPU memory leak when used iteratively
buggood first issuepriority: 1waiting on author
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
11 commentaires0 réaction2 personnes assignées
Ouverte
Make lazy initialization in plugins more robust
distributedfeaturegood first issuehelp wantedlet's do it!
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
4 commentaires0 réaction1 personne assignée
Ouverte
Allow obtaining num_nodes from ClusterEnvironment
environmentfeaturegood first issuehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
8 commentaires0 réaction0 personne assignée
Ouverte
Add checks for model spec and matching output values in `to_onnx()` method
featuregood first issuehelp wanted
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
6 commentaires2 réactions2 personnes assignées
Ouverte
CI Testing ROCm
cifeaturehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
3 commentaires4 réactions0 personne assignée
Ouverte
DDP with 2 GPUs doesn't give same results as 1 GPU with the same effective batch size
bughelp wantedpriority: 2strategy: ddpwon't fix
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
41 commentaires9 réactions0 personne assignée
Ouverte
MLFlow Logger Makes a New Run When Resuming from hpc Checkpoint
bugcheckpointingenvironment: slurmhelp wantedlogger: mlflowpriority: 2won't fix
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
8 commentaires0 réaction0 personne assignée
Ouverte
support len(datamodule)
data handlingfeaturegood first issuehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
14 commentaires0 réaction0 personne assignée
Ouverte
CLI for inspecting checkpoints
checkpointingdesignfeaturehelp wanted
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
2 commentaires7 réactions1 personne assignée
Ouverte
Load callback states while testing.
checkpointingfeaturehelp wantedpriority: 1trainer: testtrainer: validate
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
21 commentaires0 réaction0 personne assignée
Ouverte
Store logger experiment id in checkpoint to enable correct resuming of experiments
checkpointingfeaturehelp wantedlogger
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
8 commentaires8 réactions1 personne assignée
Ouverte
Resuming should allow to differentiate what to resume (steps/opti/weights)
featurehelp wantedpriority: 1
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
26 commentaires18 réactions0 personne assignée
Ouverte
Returning None from training_step with multi GPU DDP training
distributedfeaturehelp wantedpriority: 1
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
26 commentaires4 réactions1 personne assignée
Ouverte
`lr_finder` fails when called after training for 1 or more epochs
bughelp wantedpriority: 1tuner
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
15 commentaires0 réaction0 personne assignée
Ouverte
Clarify the model checkpoint arguments
callback: model checkpointhelp wantedrefactor
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
16 commentaires8 réactions0 personne assignée
Ouverte
Easier change optimizer/learning rate instead of reading state_dict from checkpoint
featurehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
12 commentaires0 réaction0 personne assignée
Ouverte
Support uneven DDP inputs with pytorch model.join
3rd partydistributedfeaturehelp wanted
Pourquoi recommandéeAucune personne assignée · Label adapté aux débutants
Aucune personne assignéeLabel adapté aux débutantsGuide de contribution disponible
27 commentaires13 réactions0 personne assignée
Ouverte
Model Verification in Trainer
featurehelp wantedlet's do it!
Pourquoi recommandéeLabel adapté aux débutants · Guide de contribution disponible
Label adapté aux débutantsGuide de contribution disponible
19 commentaires2 réactions1 personne assignée