Cedar: Difference between revisions
No edit summary |
(introduce new name + minor cleanup) |
||
Line 3: | Line 3: | ||
<translate> | <translate> | ||
</noinclude> | </noinclude> | ||
===GP2 | ===Cedar (GP2)=== <!--T:1--> | ||
<!--T:2--> | |||
CEDAR is a heterogeneous cluster, suitable for a variety of workloads, located at Simon Fraser University. It is named for the [https://en.wikipedia.org/wiki/Thuja_plicata Western Red Cedar], B.C.’s official tree and of great spiritual significance to the region's First Nations people. It was previously known as "GP2" and is still identified as such in the [https://www.computecanada.ca/research-portal/accessing-resources/resource-allocation-competitions/ 2017 RAC] documentation. | |||
<!--T:3--> | <!--T:3--> | ||
System evaluation is not yet completed as of November 2016. Anticipated specifications, based on SFU's RFP and bids received, include the following. This information is '''not guaranteed''' and might not be complete. It is provided for planning purposes. | |||
====Attached Storage System==== <!--T:4--> | ====Attached Storage System==== <!--T:4--> | ||
Line 30: | Line 31: | ||
|| | || | ||
Provided by the [[National_Data_Cyberinfrastructure|NDC]].<br /> | Provided by the [[National_Data_Cyberinfrastructure|NDC]].<br /> | ||
Available to compute nodes, but not designed for parallel | Available to compute nodes, but not designed for parallel I/O workloads.<br /> | ||
|- | |- | ||
|'''High performance interconnect''' | |'''High performance interconnect''' | ||
|| | || | ||
Low-latency high-performance fabric connecting all nodes and temporary storage. <br /> | Low-latency high-performance fabric connecting all nodes and temporary storage. <br /> | ||
The design of | The design of Cedar is to support multiple simultaneous parallel jobs of at least 1024 cores in a fully non-blocking manner. Jobs larger than 1024 cores would be less well-suited for the topology. | ||
|} | |} | ||
Line 65: | Line 66: | ||
<!--T:12--> | <!--T:12--> | ||
The | The completed system is expected to have over 25,000 cores, 900 nodes, and 500 GPUs. | ||
<!--T:13--> | <!--T:13--> | ||
The delivery and installation schedule is not yet known, but the procurement team has confidence that the system will be in production for the start of the allocations year on April 1, 2017. | The delivery and installation schedule is not yet known, but the procurement team has confidence that the system will be in production for the start of the allocations year on April 1, 2017. | ||
<noinclude> | <noinclude> | ||
</translate> | </translate> | ||
</noinclude> | </noinclude> |
Revision as of 21:30, 22 November 2016
Cedar (GP2)
CEDAR is a heterogeneous cluster, suitable for a variety of workloads, located at Simon Fraser University. It is named for the Western Red Cedar, B.C.’s official tree and of great spiritual significance to the region's First Nations people. It was previously known as "GP2" and is still identified as such in the 2017 RAC documentation.
System evaluation is not yet completed as of November 2016. Anticipated specifications, based on SFU's RFP and bids received, include the following. This information is not guaranteed and might not be complete. It is provided for planning purposes.
Attached Storage System
$HOME |
Standard home directory |
$SCRATCH Parallel High-performance filesystem |
Approximately 4PB usable capacity for temporary ( |
$PROJECT External persistent storage |
Provided by the NDC. |
High performance interconnect |
Low-latency high-performance fabric connecting all nodes and temporary storage. |
Node types and characteristics:
"Base" compute nodes: | Over 500 nodes | 128 GB of memory, 16 cores/socket, 2 sockets/node |
"Large" compute nodes: | Over 100 nodes | 256 GB of memory, 16 cores/socket, 2 sockets/node |
"Bigmem512" | 24 nodes | 512 GB of memory, 16 cores/socket, 2 sockets/node |
"Bigmem1500" nodes | 24 nodes | 1.5 TB of memory, 16 cores/socket, 2 sockets/node |
"GPU base" nodes: | Over 100 nodes | 128 GB of memory, 12 cores/socket, 2 sockets/node, 4 GPUs/node, 2 GPUs/socket. |
"GPU large" nodes. | Approximately 30 nodes | 256 GB of memory, 12 cores/socket, 2 sockets/node, 4 GPUs/node, All GPUs on the same socket |
"Bigmem3000" nodes | 4 nodes | 3 TB of memory, 8 cores/socket, 4 sockets/node |
All of the above nodes will have local (on-node) storage.
Compute Canada is not currently able to disclose the specific GPU model or specifications. The RFP used NVIDIA K80 as a baseline specification.
The completed system is expected to have over 25,000 cores, 900 nodes, and 500 GPUs.
The delivery and installation schedule is not yet known, but the procurement team has confidence that the system will be in production for the start of the allocations year on April 1, 2017.