Back to Table of contents

Primeur weekly 2018-10-08

Special

Where did the first 500 million euro invested by the European Horizon 2020 programme go? ...

Focus

World's first ARM-based supercomputer Isambard is ready for science ...

Exascale supercomputing

New European project ESCAPE-2 on exascale computing for numerical weather prediction gets under way ...

Berkeley Lab, Oak Ridge, and NVIDIA team breaks exaop barrier with deep learning application ...

Coming soon to exascale computing: Software for chemistry of catalysis ...

Quantum computing

ORNL researchers advance quantum computing, science through six DOE awards ...

Berkeley Lab to build an advanced quantum computing testbed ...

Berkeley Lab to push quantum information frontiers with new programmes in computing, physics, materials, and chemistry ...

Berkeley Quantum to accelerate innovation in quantum information science ...

Quantum software company Zapata Computing adds Clark Golestani to Board ...

Defects promise quantum communication through standard optical fiber ...

Focus on Europe

Atos and the University of Reims launch ROMEO, one of the most powerful supercomputers in the world, under the sponsorship of Cedric Villani ...

Special Edition of Open e-IRG Workshop under the Austrian EU Presidency will focus on relationship between Open Science, FAIR data and EOSC ...

Goethe University to develop green supercomputer for science ...

Calling on HPC experts and enthusiasts to propose tutorials and workshops for ISC 2019 ...

ISC 2019 calls for research paper submission by December 12, 2018 ...

Middleware

USC ISI to pilot Cyberinfrastructure Center of Excellence for National Science Foundation ...

Hardware

Tintri co-founder Mark Gritter joins Tintri by DDN as CTO to lead analytics and server virtualization vision ...

DDN simplifies the AI data centre with NVIDIA ...

New research could lead to more energy-efficient computing ...

Applications

New simulation sheds light on spiraling supermassive black holes ...

DNA unzipped, turned around, and rezipped ...

Dark Energy Survey releases first year value-added data products ...

A quantum leap toward expanding the search for dark matter ...

HP-CONCORD paves the way for scalable machine learning in HPC ...

In disaster's wake, novel computing techniques support emergency responders ...

Transition metal dichalcogenides could increase computer speed, memory by a million times ...

A new brain-inspired architecture could improve how computers handle data and advance AI ...

Rochester Institute of Technology leads multi-university collaboration to simulate neutron star mergers ...

The Cloud

Oracle rolls out Autonomous NoSQL Database service ...

Quanta Cloud Technology showcases AI portfolio options at GTC Europe ...

ZeroStack delivers GPU-as-a-Service via NVIDIA hardware ...

Berkeley Lab, Oak Ridge, and NVIDIA team breaks exaop barrier with deep learning application

High-quality segmentation results produced by deep learning on climate datasets.5 Oct 2018 Berkeley - A team of computational scientists from Lawrence Berkeley National Laboratory and Oak Ridge National Laboratory (ORNL) and engineers from NVIDIA has, for the first time, demonstrated an exascale-class deep learning application that has broken the exaop barrier.

Using a climate dataset from Berkeley Lab on ORNL's Summit system at the Oak Ridge Leadership Computing Facility (OLCF), they trained a deep neural network to identify extreme weather patterns from high-resolution climate simulations. Summit is an IBM Power Systems AC922 system powered by more than 9,000 IBM POWER9 CPUs and 27,000 NVIDIA Tesla V100 Tensor Core GPUs. By tapping into the specialized NVIDIA Tensor Cores built into the GPUs at scale, the researchers achieved a peak performance of 1.13 exaops and a sustained performance of 0.999 - the fastest deep learning algorithm reported to date and an achievement that earned them a spot on this year's list of finalists for the Gordon Bell Prize.

"This collaboration has produced a number of unique accomplishments", stated Prabhat, who leads the Data & Analytics Services team at Berkeley Lab's National Energy Research Scientific Computing Center and is a co-author on the Gordon Bell submission. "It is the first example of deep learning architecture that has been able to solve segmentation problems in climate science, and in the field of deep learning, it is the first example of a real application that has broken the exascale barrier."

These achievements were made possible through an innovative blend of hardware and software capabilities. On the hardware side, Summit has been designed to deliver 200 petaflops of high-precision computing performance and was recently named the fastest computer in the world, capable of performing more than three exaops (3 billion billion calculations) per second. The system features a hybrid architecture; each of its 4,608 compute nodes contains two IBM POWER9 CPUs and six NVIDIA Volta Tensor Core GPUs, all connected via the NVIDIA NVLink high-speed interconnect.The NVIDIA GPUs are a key factor in Summit's performance, enabling up to 12 times higher peak teraflops for training and 6 times higher peak teraflops for inference in deep learning applications compared to its predecessor, the Tesla P100.

"Our partnering with Berkeley Lab and Oak Ridge National Laboratory showed the true potential of NVIDIA Tensor Core GPUs for AI and HPC applications", stated Michael Houston, senior distinguished engineer of deep learning at NVIDIA. "To make exascale a reality, our team tapped into the multi-precision capabilities packed into the thousands of NVIDIA Volta Tensor Core GPUs on Summit to achieve peak performance in training and inference in deep learning applications."

On the software side, in addition to providing the climate dataset, the Berkeley Lab team developed pattern-recognition algorithms for training the DeepLabv3+ neural network to extract pixel-level classifications of extreme weather patterns, which could aid in the prediction of how extreme weather events are changing as the climate warms. According to Thorsten Kurth, an application performance specialist at NERSC who led this project, the team made modifications to DeepLabv3+ that improved the network's scalability and communications capabilities and made the exaops achievement possible. This included tweaking the network to train it to extract pixel-level features and per-pixel classification and improve node-to-node communication.

"What is impressive about this effort is that we could scale a high-productivity framework like TensorFlow, which is technically designed for rapid prototyping on small to medium scales, to 4,560 nodes on Summit", he stated. "With a number of performance enhancements, we were able to get the framework to run on nearly the entire supercomputer and achieve exaop-level performance, which to my knowledge is the best achieved so far in a tightly coupled application."

Other innovations included high-speed parallel data staging, an optimized data ingestion pipeline and multi-channel segmentation. Traditional image segmentation tasks work on three-channel red/blue/green images. But scientific datasets often comprise many channels; in climate, for example, these can include temperature, wind speeds, pressure values and humidity. By running the optimized neural network on Summit, the additional computational capabilities allowed the use of all 16 available channels, which dramatically improved the accuracy of the models.

"We have shown that we can apply deep-learning methods for pixel-level segmentation on climate data, and potentially on other scientific domains", stated Prabhat. "More generally, our project has laid the groundwork for exascale deep learning for science, as well as commercial applications."

In addition to Prabhat, Michael Houston and Thorsten Kurth, the research team included Jack Deslippe, Mayur Mudigonda and Ankur Mahesh of Berkeley Lab; Michael Matheson of ORNL; and Sean Treichler, Joshua Romero, Nathan Luehr, Everett Phillips and Massimiliano Fatica of NVIDIA.

Established more than three decades ago by the Association for Computing Machinery, the Gordon Bell award recognizes outstanding achievement in the field of computing for applications in science, engineering and large-scale data science. This year's winner will be announced at SC18 in November in Dallas, Texas.

NERSC and OLCF are both DOE Office of Science User Facilities.
Source: National Energy Research Scientific Computing Center - NERSC

Back to Table of contents

Primeur weekly 2018-10-08

Special

Where did the first 500 million euro invested by the European Horizon 2020 programme go? ...

Focus

World's first ARM-based supercomputer Isambard is ready for science ...

Exascale supercomputing

New European project ESCAPE-2 on exascale computing for numerical weather prediction gets under way ...

Berkeley Lab, Oak Ridge, and NVIDIA team breaks exaop barrier with deep learning application ...

Coming soon to exascale computing: Software for chemistry of catalysis ...

Quantum computing

ORNL researchers advance quantum computing, science through six DOE awards ...

Berkeley Lab to build an advanced quantum computing testbed ...

Berkeley Lab to push quantum information frontiers with new programmes in computing, physics, materials, and chemistry ...

Berkeley Quantum to accelerate innovation in quantum information science ...

Quantum software company Zapata Computing adds Clark Golestani to Board ...

Defects promise quantum communication through standard optical fiber ...

Focus on Europe

Atos and the University of Reims launch ROMEO, one of the most powerful supercomputers in the world, under the sponsorship of Cedric Villani ...

Special Edition of Open e-IRG Workshop under the Austrian EU Presidency will focus on relationship between Open Science, FAIR data and EOSC ...

Goethe University to develop green supercomputer for science ...

Calling on HPC experts and enthusiasts to propose tutorials and workshops for ISC 2019 ...

ISC 2019 calls for research paper submission by December 12, 2018 ...

Middleware

USC ISI to pilot Cyberinfrastructure Center of Excellence for National Science Foundation ...

Hardware

Tintri co-founder Mark Gritter joins Tintri by DDN as CTO to lead analytics and server virtualization vision ...

DDN simplifies the AI data centre with NVIDIA ...

New research could lead to more energy-efficient computing ...

Applications

New simulation sheds light on spiraling supermassive black holes ...

DNA unzipped, turned around, and rezipped ...

Dark Energy Survey releases first year value-added data products ...

A quantum leap toward expanding the search for dark matter ...

HP-CONCORD paves the way for scalable machine learning in HPC ...

In disaster's wake, novel computing techniques support emergency responders ...

Transition metal dichalcogenides could increase computer speed, memory by a million times ...

A new brain-inspired architecture could improve how computers handle data and advance AI ...

Rochester Institute of Technology leads multi-university collaboration to simulate neutron star mergers ...

The Cloud

Oracle rolls out Autonomous NoSQL Database service ...

Quanta Cloud Technology showcases AI portfolio options at GTC Europe ...

ZeroStack delivers GPU-as-a-Service via NVIDIA hardware ...