Files
pytorch-lightning/examples/multi_node_examples
2019-10-05 14:30:12 -04:00
..
2019-10-05 14:13:32 -04:00
2019-10-05 14:28:08 -04:00
2019-10-05 14:21:12 -04:00
2019-10-05 14:30:12 -04:00

Multi-node example

To run this demo which launches a single job that trains on 2 nodes (2 gpus per node), do the following:

  1. Log into the jumphost node of your SLURM-managed cluster.
  2. Create a conda environment with Lightning and a GPU PyTorch version.
  3. Submit this script.
sbatch job_submit.sh your_env_name_with_lightning_installed