CERN Project Sees Orders-of-Magnitude Speedup with AI Approach

By Rob Farber

August 14, 2018

An award-winning effort at CERN has demonstrated potential to significantly change how the physics based modeling and simulation communities view machine learning. The CERN team demonstrated that AI-based models have the potential to act as orders-of-magnitude-faster replacements for computationally expensive tasks in simulation, while maintaining a remarkable level of accuracy.

Dr. Federico Carminati (Project Coordinator, CERN) points out, “This work demonstrates the potential of ‘black box’ machine-learning models in physics-based simulations.”

A poster describing this work was awarded the prize for best poster in the category ‘programming models and systems software’ at ISC’18. This recognizes the importance of the work, which was carried out by Dr. Federico Carminati, Gul Rukh Khattak, and Dr. Sofia Vallecorsa at CERN, as well as Jean-Roch Vlimant at Caltech. The work is part of a CERN openlab project in collaboration with Intel Corporation, who partially funded the endeavor through the Intel Parallel Computing Center (IPCC) program.

Widespread potential impact for simulation

The world-wide impact for High-Energy Physics (HEP) scientists could be substantial, as outlined by the CERN poster, which points out that ”Currently, most of the LHC’s worldwide distributed CPU budget — in the range of half a million CPU-years equivalent — is dedicated to simulation.” Speeding up the most time-consuming simulation tasks (e.g., high-granularity calorimeters, which are components in a detector that measure the energy of particles[i]) will help scientists better utilize these allocations. The following are comparative results obtained by the CERN team in the time to create an electron shower, once the AI model has been fully trained:

Figure 1: Comparative runtime to create an electron shower of the machine-learning method (e.g. 3d GAN) vs. the full Monte-Carlo simulation (Image courtesy CERN)

Dr. Sofia Vallecorsa points out that the CPU based runtime is important as nearly all of the Geant user base runs on CPUs. Vallecorsa is a CERN physicist who was also highlighted in the CERN article Coding has no gender.

As scientists consider future CERN experiments, Vallecorsa observes, “Given future plans to upgrade CERN’s Large Hadron Collider, dramatically increasing particle collision rates, frameworks like this have the potential to play an important role in ensuring data rates remain manageable.”

This kind of approach could help to realize similar orders-of-magnitude-faster speedups for computationally expensive simulation tasks used in a range of fields.

Vallecorsa explains that the data distributions coming from the trained machine-learning model are remarkably close to the real and simulated data.

A big change in thinking

The team demonstrated that “energy showers” detected by calorimeters can be interpreted as a 3D image[ii]. The process is illustrated in the following figure. The team adopted this approach from the machine-learning community as deep-learning convolutional neural networks are heavily utilized when working with images.

Figure 2: Schematic from the poster showing how a single particle creates an electron shower that can be viewed as an image (Courtesy CERN)

Use of GANS

The CERN team decided to train Generative Adversarial Networks (GANs) on the calorimeter images. GANs are particularly suited to act as a replacement for the expensive Monte Carlo methods used in HEP simulations as they generate realistic samples for complicated probability distributions, allow multi-modal output, can do interpolation, and are robust against missing data.

The basic idea is easy to understand: train a Generator (G) to create the calorimeter image with sufficient accuracy to trick a discriminator (D) which tries to identify artificial samples from the generator compared to real samples from the Monte Carlo simulation. G reproduces the data distribution starting from random noise. D estimates the probability that a sample came from the training data rather than G. The training procedure for G is to maximize the probability of D making a mistake. A high-level illustration of the GAN is provided below.

Figure 3: High-level view of training a GAN (image from https://medium.com/@devnag/generative-adversearial-networks-in-50-lines-of-code-pytorch-e81b79659e3f)

Even though the description is simple, 3D GANs are unfortunately not “out-of-the-box” networks, which meant the training of the model was non-trivial.

Results

After detailed validation of the trained GAN, there was “remarkable” agreement between the images from the generator and the Monte-Carlo images. This type of approach could potentially be beneficial in other fields where Monte Carlo simulation is used.

More specifically, the CERN team compared high level quantities (e.g., energy shower shapes) and detailed calorimeter response (e.g., single cell response) between the trained generator and the standard Monte Carlo. The CERN team describes the agreement, which is within a few percent, as “remarkable” in their poster.

Visually this agreement can be seen by how closely the blue (real data) and red lines (GAN generated data) overlap in the following results reported in the poster.

Figure 4: Transverse shower shape for 100-500 GeV pions. Red is the GAN data while blue represents the real data. (Image courtesy CERN)

 

Figure 5: Longitudinal shower shape for 400 GeV electron (Image courtesy CERN)

 

Figure 6: Longitudinal shower shape for 100 GeV electron (Image courtesy CERN)

Vallecorsa summarizes these results by stating, “The agreement between the images generated by our model and the Monte Carlo images has been beyond our expectations. This demonstrates that this is a promising avenue for further investigation.”

CERN openlab

The CERN team plans to test performance using FPGAs and other integrated accelerator technologies. FPGAs are known to deliver lower latency and higher inferencing performance than both CPUs and GPUs[iii]. The CERN group also intends to test several deep learning techniques in the hope of achieving a yet greater speedup with respect to Monte Carlo techniques, and ensuring this approach covers a range of detector types, which CERN believes is key to future projects.

This research is being carried out through a CERN openlab project. CERN openlab is a public-private partnership through which CERN collaborates with leading ICT companies to drive innovation in cutting-edge ICT solutions for its research community. Intel has been a partner in CERN openlab since it was first established in 2001. Dr. Alberto Di Meglio (Head of CERN openlab) observes, “At CERN, we’re always interested in exploring upcoming technologies that can help researchers to make new ground-breaking discoveries about our universe. We support this through joint R&D projects with our collaborators from industry, and by making cutting-edge technologies available for evaluation by researchers at CERN.”

Summary

The HPC modeling and simulation community now has a promising path forward to exploit the benefits of machine learning. The key, as demonstrated by CERN, is that the machine-learning-generated distribution needs to be indistinguishable from other high-fidelity methods in physics-based simulations.

The motivation is straightforward: (1) orders of magnitude faster performance, (2) efficient CPU implementations, and (3) this approach could enable the use of other new technologies such as FPGAs that may significantly improve performance.

Additional References

Rob Farber is a global technology consultant and author with an extensive background in HPC and in machine learning technology that he applies at national labs and commercial organizations on a variety of problems including challenges in high energy physics. Rob can be reached at [email protected]

[i] http://cds.cern.ch/record/2254048#

[ii] ibid

[iii] https://medium.com/syncedreview/deep-learning-in-real-time-inference-acceleration-and-continuous-training-17dac9438b0b

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industry updates delivered to you every week!

MLPerf Inference 4.0 Results Showcase GenAI; Nvidia Still Dominates

March 28, 2024

There were no startling surprises in the latest MLPerf Inference benchmark (4.0) results released yesterday. Two new workloads — Llama 2 and Stable Diffusion XL — were added to the benchmark suite as MLPerf continues Read more…

Q&A with Nvidia’s Chief of DGX Systems on the DGX-GB200 Rack-scale System

March 27, 2024

Pictures of Nvidia's new flagship mega-server, the DGX GB200, on the GTC show floor got favorable reactions on social media for the sheer amount of computing power it brings to artificial intelligence.  Nvidia's DGX Read more…

Call for Participation in Workshop on Potential NSF CISE Quantum Initiative

March 26, 2024

Editor’s Note: Next month there will be a workshop to discuss what a quantum initiative led by NSF’s Computer, Information Science and Engineering (CISE) directorate could entail. The details are posted below in a Ca Read more…

Waseda U. Researchers Reports New Quantum Algorithm for Speeding Optimization

March 25, 2024

Optimization problems cover a wide range of applications and are often cited as good candidates for quantum computing. However, the execution time for constrained combinatorial optimization applications on quantum device Read more…

NVLink: Faster Interconnects and Switches to Help Relieve Data Bottlenecks

March 25, 2024

Nvidia’s new Blackwell architecture may have stolen the show this week at the GPU Technology Conference in San Jose, California. But an emerging bottleneck at the network layer threatens to make bigger and brawnier pro Read more…

Who is David Blackwell?

March 22, 2024

During GTC24, co-founder and president of NVIDIA Jensen Huang unveiled the Blackwell GPU. This GPU itself is heavily optimized for AI work, boasting 192GB of HBM3E memory as well as the the ability to train 1 trillion pa Read more…

MLPerf Inference 4.0 Results Showcase GenAI; Nvidia Still Dominates

March 28, 2024

There were no startling surprises in the latest MLPerf Inference benchmark (4.0) results released yesterday. Two new workloads — Llama 2 and Stable Diffusion Read more…

Q&A with Nvidia’s Chief of DGX Systems on the DGX-GB200 Rack-scale System

March 27, 2024

Pictures of Nvidia's new flagship mega-server, the DGX GB200, on the GTC show floor got favorable reactions on social media for the sheer amount of computing po Read more…

NVLink: Faster Interconnects and Switches to Help Relieve Data Bottlenecks

March 25, 2024

Nvidia’s new Blackwell architecture may have stolen the show this week at the GPU Technology Conference in San Jose, California. But an emerging bottleneck at Read more…

Who is David Blackwell?

March 22, 2024

During GTC24, co-founder and president of NVIDIA Jensen Huang unveiled the Blackwell GPU. This GPU itself is heavily optimized for AI work, boasting 192GB of HB Read more…

Nvidia Looks to Accelerate GenAI Adoption with NIM

March 19, 2024

Today at the GPU Technology Conference, Nvidia launched a new offering aimed at helping customers quickly deploy their generative AI applications in a secure, s Read more…

The Generative AI Future Is Now, Nvidia’s Huang Says

March 19, 2024

We are in the early days of a transformative shift in how business gets done thanks to the advent of generative AI, according to Nvidia CEO and cofounder Jensen Read more…

Nvidia’s New Blackwell GPU Can Train AI Models with Trillions of Parameters

March 18, 2024

Nvidia's latest and fastest GPU, codenamed Blackwell, is here and will underpin the company's AI plans this year. The chip offers performance improvements from Read more…

Nvidia Showcases Quantum Cloud, Expanding Quantum Portfolio at GTC24

March 18, 2024

Nvidia’s barrage of quantum news at GTC24 this week includes new products, signature collaborations, and a new Nvidia Quantum Cloud for quantum developers. Wh Read more…

Alibaba Shuts Down its Quantum Computing Effort

November 30, 2023

In case you missed it, China’s e-commerce giant Alibaba has shut down its quantum computing research effort. It’s not entirely clear what drove the change. Read more…

Nvidia H100: Are 550,000 GPUs Enough for This Year?

August 17, 2023

The GPU Squeeze continues to place a premium on Nvidia H100 GPUs. In a recent Financial Times article, Nvidia reports that it expects to ship 550,000 of its lat Read more…

Shutterstock 1285747942

AMD’s Horsepower-packed MI300X GPU Beats Nvidia’s Upcoming H200

December 7, 2023

AMD and Nvidia are locked in an AI performance battle – much like the gaming GPU performance clash the companies have waged for decades. AMD has claimed it Read more…

DoD Takes a Long View of Quantum Computing

December 19, 2023

Given the large sums tied to expensive weapon systems – think $100-million-plus per F-35 fighter – it’s easy to forget the U.S. Department of Defense is a Read more…

Synopsys Eats Ansys: Does HPC Get Indigestion?

February 8, 2024

Recently, it was announced that Synopsys is buying HPC tool developer Ansys. Started in Pittsburgh, Pa., in 1970 as Swanson Analysis Systems, Inc. (SASI) by John Swanson (and eventually renamed), Ansys serves the CAE (Computer Aided Engineering)/multiphysics engineering simulation market. Read more…

Choosing the Right GPU for LLM Inference and Training

December 11, 2023

Accelerating the training and inference processes of deep learning models is crucial for unleashing their true potential and NVIDIA GPUs have emerged as a game- Read more…

Intel’s Server and PC Chip Development Will Blur After 2025

January 15, 2024

Intel's dealing with much more than chip rivals breathing down its neck; it is simultaneously integrating a bevy of new technologies such as chiplets, artificia Read more…

Baidu Exits Quantum, Closely Following Alibaba’s Earlier Move

January 5, 2024

Reuters reported this week that Baidu, China’s giant e-commerce and services provider, is exiting the quantum computing development arena. Reuters reported � Read more…

Leading Solution Providers

Contributors

Comparing NVIDIA A100 and NVIDIA L40S: Which GPU is Ideal for AI and Graphics-Intensive Workloads?

October 30, 2023

With long lead times for the NVIDIA H100 and A100 GPUs, many organizations are looking at the new NVIDIA L40S GPU, which it’s a new GPU optimized for AI and g Read more…

Shutterstock 1179408610

Google Addresses the Mysteries of Its Hypercomputer 

December 28, 2023

When Google launched its Hypercomputer earlier this month (December 2023), the first reaction was, "Say what?" It turns out that the Hypercomputer is Google's t Read more…

AMD MI3000A

How AMD May Get Across the CUDA Moat

October 5, 2023

When discussing GenAI, the term "GPU" almost always enters the conversation and the topic often moves toward performance and access. Interestingly, the word "GPU" is assumed to mean "Nvidia" products. (As an aside, the popular Nvidia hardware used in GenAI are not technically... Read more…

Shutterstock 1606064203

Meta’s Zuckerberg Puts Its AI Future in the Hands of 600,000 GPUs

January 25, 2024

In under two minutes, Meta's CEO, Mark Zuckerberg, laid out the company's AI plans, which included a plan to build an artificial intelligence system with the eq Read more…

Google Introduces ‘Hypercomputer’ to Its AI Infrastructure

December 11, 2023

Google ran out of monikers to describe its new AI system released on December 7. Supercomputer perhaps wasn't an apt description, so it settled on Hypercomputer Read more…

China Is All In on a RISC-V Future

January 8, 2024

The state of RISC-V in China was discussed in a recent report released by the Jamestown Foundation, a Washington, D.C.-based think tank. The report, entitled "E Read more…

Intel Won’t Have a Xeon Max Chip with New Emerald Rapids CPU

December 14, 2023

As expected, Intel officially announced its 5th generation Xeon server chips codenamed Emerald Rapids at an event in New York City, where the focus was really o Read more…

IBM Quantum Summit: Two New QPUs, Upgraded Qiskit, 10-year Roadmap and More

December 4, 2023

IBM kicks off its annual Quantum Summit today and will announce a broad range of advances including its much-anticipated 1121-qubit Condor QPU, a smaller 133-qu Read more…

  • arrow
  • Click Here for More Headlines
  • arrow
HPCwire