Files
pytorch-lightning/examples/multi_node_examples/README.md
T
2019-10-05 15:21:32 -04:00

336 B

Multi-node example

To run this demo which launches a single job that trains on 2 nodes (2 gpus per node), do the following:

  1. Log into the jumphost node of your SLURM-managed cluster.
  2. Create a conda environment with Lightning and a GPU PyTorch version.
  3. Submit this script.
sbatch job_submit.sh YourEnv