Content Comparison

...

40-core 96GB
40-core 192GB
56-core 128GB
56-core 256GB
56-core 512GB
64-core 192GB
64-core 384GB
64-core 768GB
80-core 96GB
80-core 192GB
80-core 384GB
80-core 768GB
80-core 1.5TB
112-core 256GB
112-core 512GB
112-core 1024GB
112-core 1.5TB
128-core 256GB
128-core 512GB
128-core 1TB
128-core 1.5TB

...

21 machines with Nvidia P100 accelerators
2 machines with Nvidia K80 accelerators
2 machines with Nvidia P40 accelerators
17 machines with 1080Ti accelerators
19 machines with Titan V accelerators
14 machines with V100 accelerators
38 machines with 2080Ti accelerators
1 machine with RTX8000 accelerators
7 machines with A100 accelerators
5 machines with 4 A40 accelerators each
2 machines with 4 L40S accelerators each
1 machine with 4 L4 accelerators

Heterogeneity

...

Table plus

Node memory (GB)	Job slots	Memory (GB) per slot
96	40	2
96	80	1
128	56	2
192	40	5
192	64	3
192	80	2
256	56	4
256	112	2
256	128	2
384	64	6
384	80	5
512	56	9
512	112	4
512	128	4
768	64	12
768	80	9
1024	112	9
1024	128	8
1536	80	19
1536	112	13
1536	128	12

Using the Basic Job Submission and Advanced Job Submission pages as a reference, how would one submit jobs taking HT into account? For single process high throughput type jobs it probably does not matter, just request one slot per job. For multithreaded or MPI jobs, request one job slot per thread or process. So if your application runs best with 4 threads then request something like the following.

...

Version	Old Version 219	New Version Current
Changes made by	John Saxton	John Saxton
Saved on	Jul 10, 2024	Nov 22, 2024

Versions Compared

Key

Heterogeneity