Hey there 👋🏽, have you ever wondered how to run deep learning workflows on a GPU cluster? Welcome to our in-depth tutorial! We’ll walk you through the process of deploying an example workflow onto a GPU cluster, from initial logging into the cluster right through to monitoring your training process. To access the cluster, you’ll need an NHR account (get it here: https://zulassung.hlrn.de), available for free to all researchers based in Germany.
This video covers how to execute code on the compute nodes with the Slurm command sbatch.
Resources
📚 Course material: https://gitlab-ce.gwdg.de/dmuelle3/de...
🤔 First time on a cluster? Get to know the concepts https://gitlab-ce.gwdg.de/hpc-team-pu...
👾 GPU documentation on NHR: https://www.hlrn.de/doc/display/PUB/G...
⏱️ Slurm scheduler documentation: https://slurm.schedmd.com/documentati...
Chapters
00:00 Intro
00:48 Example script
02:07 Live demonstration sbatch
04:29 Check status (squeue)
05:05 Cancel jobs (scancel)
Sur cette page du site, vous pouvez voir la vidéo en ligne 3.3 Slurm: Run code with sbatch [Deep Learning + GPU Tutorial] durée heure minute seconde en bonne qualité , qui a été Téléchargé par l'utilisateur GWDG 07 juillet 2023, Partagez le lien avec vos amis et connaissances, sur youtube cette vidéo a déjà été regardée 4,496 fois et il a aimé 63 téléspectateurs. Bon visionnage!