The University of Edinburgh is upgrading its Cirrus supercomputer with 144 specialised GPU accelerators to help scientists prepare for the next generation of computing. Most current supercomputers rely on standard processors (CPUs), but the first exascale machines—capable of a billion billion calculations per second—are expected by 2020 or 2021, and they may use a hybrid CPU-GPU design. This creates a problem: researchers do not yet know how well their simulation software will run on such heterogeneous systems. Cirrus Phase II gives UK scientists a testbed to find out. The system also provides a high-performance platform for artificial intelligence training, which is increasingly important in both scientific and industrial applications. If the project succeeds, it will help ensure that UK researchers can make effective use of future exascale computers, keeping fields like climate modelling, drug discovery, and materials design competitive. The work is primarily about preparing computing infrastructure for a technological shift, rather than delivering immediate real-world applications.
View original technical description
EPSRC's Tier 2 High Performance Computing (HPC) services provide access to supercomputing systems that complement the national HPC service, ARCHER. In 2016 a national network of five Tier 2 HPC services was established across the UK. One of these services, Cirrus, is hosted and operated by EPCC, the supercomputing centre at the University of Edinburgh. The Cirrus service has been very successful, currently with over 380 active users of which over 50 are from industry. The Cirrus Phase II service will build on the success of the Cirrus service by adding tightly integrated GPU accelerators to the system. GPU accelerators help supercomputing applications run faster by accelerating their core numerical calculations. Cirrus Phase II will have 144 NVIDIA V100 GPUs and an accompanying fast storage layer for data intensive applications. Systems that include both normal CPUs and GPU accelerators are often called heterogenous systems. The next frontier of supercomputing is the Exascale and the first Exascale systems will become operational in 2020 or 2021. There are two different technology solutions which can delvier an Exaflop (a billion billion calculations per second). One option is a CPU only approach the other approach is CPU and GPU heterogeneous approach. The enhanced Cirrus Phase II system will give scientists from across the UK the opportunity to explore how their modelling and simulation applications will run on a heterogeneous architecture and how difficult it is likely to be to deliver sufficient performance at scale. In addition to exploring Exascale options, the system will also provide a high performance platform for AI training and research. Currently there is enormous interest in using AI in scientific and industrial applications and Cirrus Phase II, with its large number of NVIDIA GPUs and fast storage layer will provide an excellent platform for such applications.
Plain English summaries and category classifications on this site are generated by AI and may not perfectly reflect the original research.
Is something wrong? Let us know