Filippo Mantovani on What’s Next for Mont-Blanc and ARM

By John Russell

July 6, 2015

Firing up the Mont-Blanc prototype in mid-June at the Barcelona Supercomputing Center (BSC) was a significant milestone in the European effort to base HPC systems on energy efficient architecture. Mont-Blanc program coordinator Filippo Mantovani was quoted in the release announcing the prototype saying, “Now the challenge starts because with this platform we can foresee how inexpensive technologies from the mobile market can be leveraged for traditional scientific high-performance workloads.”

Begun in 2011, the Mont-Blanc Project is European effort intended to explore new ways to achieve energy efficient architecture for supercomputing (See the 2013 Mont-Blanc paper, Supercomputing with Commodity CPUs: Are Mobile SoCs Ready for HPC?”). Recently, the project received a three extension to further develop the OmpSs parallel programming model to automatically exploit multiple cluster nodes, transparent application check pointing for fault tolerance, support for ARMv8 64-bit processors, and the initial design of the Mont-Blanc exascale architecture.

The prototype installed in the Torre Girona chapel is made up of a total of two racks containing 8 standard BullX chassis, 72 compute blades fitting 1080 compute cards, for a total of 2160 CPUs and 1080 GPUs. The heterogeneous architecture of the Mont-Blanc prototype takes advantage of computing elements (CPUs and GPUs) developed by ARM and integrated by BULL under the design guidance of all Mont-Blanc partners.

This use of the ARM architecture is an early demonstration that it may have applicability at the high end of computing. HPCwire talked with Mantovani about some of the challenges and promise that lie ahead.

What further enhancements to the ARM architecture are needed to maintain progress towards higher performance and what changes do you expect over the next few years?

It depends which ARM processors are we looking at. Enhancements of mobile System on Chips (SoCs) are driven by big producers of mobile devices (Apple, Samsung, Huawei, etc.). From this market we will see surprisingly good and increasingly powerful SoCs, but I consider unlikely that one of them will be integrated as-is in a high-end HPC system, unless some of these big players want to enter HPC market. Due to its cost effectiveness, I [still] consider [that] mobile technology is extremely interesting for compute intensive embedded applications as well as small labs and companies looking for cheap/mobile/easy scientific computation, not necessarily in the HPC area.

If we are looking at ARM processors in the server market, then the things are slightly different. The ARM-based chips for servers, in fact, seem to evolve fast and [are becoming] more popular (X-Gene, Cavium ThunderX). Strangely enough, I consider it more urgent to have reliable and unified software support for the ARM platforms appearing on the market, than adding specific features to the silicon. This support would allow ARM technology to be “better socially accepted” within the HPC community. In this sense, Mont-Blanc is going to contribute with this system software stack and programming model, but in terms of compilers a strong contribution from IP designers and SoC producers is [still] required.

What are the missing or weaker parts of the HPC ecosystem required to support continued progress of the ARM-based architecture approach? How are those pieces likely to be developed or strengthened?

Decoupling the production of HPC solutions among IP providers, SoC producers and system integrators can increase competitiveness with benefits for the diversification of solutions and prices; but it can also drive to fragmentation. HPC system integrators are mostly conservative: they are definitely not used to working with mobile technology and also ARM-based server solutions are still not 100% in the production lines of big HPC players. We saw some interesting movement during last SC in New Orleans and I really hope to see even more activity in this direction soon at ISC in Frankfurt.

I think that the real difference could be done now by a good, large, stable and most importantly open-source software support to the ARM-architecture, especially for HPC. I am thinking of compilers, support for hardware counters, parallel debuggers, performance analysis tools, etc. but also programming models that can support the proliferation of threads, the heterogeneity and the different ARMv8 implementations appearing on the market. In this sense, Mont-Blanc is doing a huge effort porting and promoting not only the development tools, but also the OmpSs programming model.

Given the prospect of reduced cost – power and hardware – do you expect ARM-based HPC to further ‘democratize’ HPC and spur adoption by industry sectors and smaller companies previously unable to afford advanced compute resources?

HPC remains mostly an “elite” market. I think however that there are several companies and small labs that have HPC-like problems, looking for accessible compute solutions. In this sense, yes, I believe that ARM-based scientific computation has a great potential. You ask for adopters? I do not have a crystal ball, but I see automotive as a potential growing market. Another field that could take advantage of cost-effective solutions could be personalized medicine. As I said, I see the potential, but I do not know how fast each of these communities reacts to new technologies appearing on the market.

Maybe less directly profitable, but I think we should not ignore the educational impact of parallel ARM-based platforms. Parallela is a worldwide example, but I think that also the fact that a team of six students will take part to the “Student Cluster Competition” at ISC’15 for the first time in the history of the contest with an ARM-based cluster (part of the Mont-Blanc prototype) must to be taken into account. Parallel, accessible and powerful platforms will help new generation of students to grow from day-zero thinking in parallel and taking into account power limitations.

What do you see as the most significant technical problems the Mont-Blanc project must solve now to achieve the next level of performance. Will new technologies be needed to solve some of these issues?

“I think that we can still extract a significant amount of information from our “large” prototype: performance evaluation at level of compute node, at system level, at level of applications, at level of fault tolerance, at level of energy to solution and at level of programmability. We will continue studying on our unique platform, this is sure.

We will approach next level of performance exploring ARM 64-bit instruction set, mostly with platforms available on both markets, server and mobile. On the software side we will continue the exploration using a larger and more complete set of performance analysis tools and boosting our task based programming model OmpSs.”

Considering the hurdles ahead, do you think an exascale system based the ARM/GPU architecture will be built and roughly when do you think we might expect it? Will we ever see a system such as this in the Top500?

“In general, for classical HPC, I consider [the] exascale target still too blurry for giving a clear prediction. Even less, unfortunately, can I foresee concerning ARM/GPU based solutions. For sure the exascale race is wider than simply finding the right technology for floating point computations: it involves memory technology, interconnection network, distributed I/O, fault tolerance and many other hardware and software aspects. In this wider approach to next generation HPC systems, I consider ARM as one of the players with great potential.”

What were the important lessons learned from the End-User Group – Rolls Royce, for example – and how will they inform Mont-Blanc development going forward? Can you identify specific issues that will need to be addressed?

The End-User Group (EUG) is an extremely valuable dissemination tool for the project, but most importantly a virtual gate for letting companies entering the development of the project. The fixed appointments are a yearly meeting with the end-users, plus the training that the project opens to the partners and to the EUG as well.

You mentioned Rolls Royce: we had very fruitful interaction during the first year of collaboration, so we decided to invite a representative to show Rolls Royce work on one of the Mont-Blanc mini-cluster at the satellite event of the PRACEdays in Dublin. The title of the workshop was emblematic, “Enabling Exascale in Europe for Industry”, and we really wanted to leave space to one of our end-users, to understand the tests performed and listen at the requirements.

I think it has been a really productive interaction and I hope that from now on, with 1000 nodes of the Mont-Blanc prototype up and running, this can evolve further, involving several other companies interested in testing the Mont-Blanc platforms.”

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industy updates delivered to you every week!

At ISC – Goh on Go: Humans Can’t Scale, the Data-Centric Learning Machine Can

June 22, 2017

I've seen the future this week at ISC, it’s on display in prototype or Powerpoint form, and it’s going to dumbfound you. The future is an AI neural network designed to emulate and compete with the human brain. In thi Read more…

By Doug Black

How ‘Knights Mill’ Gets Its Deep Learning Flops

June 22, 2017

Intel, the subject of much speculation regarding the delayed or potentially canceled “Aurora” contract (the Argonne Lab part of the CORAL “pre-exascale” award), parsed out additional information about the upc Read more…

By Tiffany Trader

GPUs, Power9, Figure Prominently in IBM’s Bet on Weather Forecasting

June 22, 2017

IBM jumped into the weather forecasting business roughly a year and a half ago by purchasing The Weather Company. This week at ISC 2017, Big Blue rolled out plans to push deeper into climate science and develop more gran Read more…

By John Russell

Intersect 360 at ISC: HPC Industry at $44B by 2021

June 22, 2017

The care, feeding and sustained growth of the HPC industry increasingly is in the hands of the commercial market sector – in particular, it’s the hyperscale companies and their embrace of AI and deep learning – tha Read more…

By Doug Black

HPE Extreme Performance Solutions

Creating a Roadmap for HPC Innovation at ISC 2017

In an era where technological advancements are driving innovation to every sector, and powering major economic and scientific breakthroughs, high performance computing (HPC) is crucial to tackle the challenges of today and tomorrow. Read more…

AMD Charges Back into the Datacenter and HPC Workflows with EPYC Processor

June 20, 2017

AMD is charging back into the enterprise datacenter and select HPC workflows with its new EPYC 7000 processor line, code-named Naples, announced today at a “global” launch event in Austin TX. In many ways it was a fu Read more…

By John Russell

Hyperion: Deep Learning, AI Helping Drive Healthy HPC Industry Growth

June 20, 2017

To be at the ISC conference in Frankfurt this week is to experience deep immersion in deep learning. Users want to learn about it, vendors want to talk about it, analysts and journalists want to report on it. Deep learni Read more…

By Doug Black

OpenACC Shows Growing Strength at ISC

June 19, 2017

OpenACC is strutting its stuff at ISC this year touting expanding membership, a jump in downloads, favorable benchmarks across several architectures, new staff members, and new support by key HPC applications providers, Read more…

By John Russell

Top500 Results: Latest List Trends and What’s in Store

June 19, 2017

Greetings from Frankfurt and the 2017 International Supercomputing Conference where the latest Top500 list has just been revealed. Although there were no major shakeups -- China still has the top two spots locked with th Read more…

By Tiffany Trader

At ISC – Goh on Go: Humans Can’t Scale, the Data-Centric Learning Machine Can

June 22, 2017

I've seen the future this week at ISC, it’s on display in prototype or Powerpoint form, and it’s going to dumbfound you. The future is an AI neural network Read more…

By Doug Black

How ‘Knights Mill’ Gets Its Deep Learning Flops

June 22, 2017

Intel, the subject of much speculation regarding the delayed or potentially canceled “Aurora” contract (the Argonne Lab part of the CORAL “pre-exascal Read more…

By Tiffany Trader

GPUs, Power9, Figure Prominently in IBM’s Bet on Weather Forecasting

June 22, 2017

IBM jumped into the weather forecasting business roughly a year and a half ago by purchasing The Weather Company. This week at ISC 2017, Big Blue rolled out pla Read more…

By John Russell

Intersect 360 at ISC: HPC Industry at $44B by 2021

June 22, 2017

The care, feeding and sustained growth of the HPC industry increasingly is in the hands of the commercial market sector – in particular, it’s the hyperscale Read more…

By Doug Black

AMD Charges Back into the Datacenter and HPC Workflows with EPYC Processor

June 20, 2017

AMD is charging back into the enterprise datacenter and select HPC workflows with its new EPYC 7000 processor line, code-named Naples, announced today at a “g Read more…

By John Russell

Hyperion: Deep Learning, AI Helping Drive Healthy HPC Industry Growth

June 20, 2017

To be at the ISC conference in Frankfurt this week is to experience deep immersion in deep learning. Users want to learn about it, vendors want to talk about it Read more…

By Doug Black

OpenACC Shows Growing Strength at ISC

June 19, 2017

OpenACC is strutting its stuff at ISC this year touting expanding membership, a jump in downloads, favorable benchmarks across several architectures, new staff Read more…

By John Russell

Top500 Results: Latest List Trends and What’s in Store

June 19, 2017

Greetings from Frankfurt and the 2017 International Supercomputing Conference where the latest Top500 list has just been revealed. Although there were no major Read more…

By Tiffany Trader

Quantum Bits: D-Wave and VW; Google Quantum Lab; IBM Expands Access

March 21, 2017

For a technology that’s usually characterized as far off and in a distant galaxy, quantum computing has been steadily picking up steam. Just how close real-wo Read more…

By John Russell

Trump Budget Targets NIH, DOE, and EPA; No Mention of NSF

March 16, 2017

President Trump’s proposed U.S. fiscal 2018 budget issued today sharply cuts science spending while bolstering military spending as he promised during the cam Read more…

By John Russell

HPC Compiler Company PathScale Seeks Life Raft

March 23, 2017

HPCwire has learned that HPC compiler company PathScale has fallen on difficult times and is asking the community for help or actively seeking a buyer for its a Read more…

By Tiffany Trader

Google Pulls Back the Covers on Its First Machine Learning Chip

April 6, 2017

This week Google released a report detailing the design and performance characteristics of the Tensor Processing Unit (TPU), its custom ASIC for the inference Read more…

By Tiffany Trader

CPU-based Visualization Positions for Exascale Supercomputing

March 16, 2017

In this contributed perspective piece, Intel’s Jim Jeffers makes the case that CPU-based visualization is now widely adopted and as such is no longer a contrarian view, but is rather an exascale requirement. Read more…

By Jim Jeffers, Principal Engineer and Engineering Leader, Intel

Nvidia Responds to Google TPU Benchmarking

April 10, 2017

Nvidia highlights strengths of its newest GPU silicon in response to Google's report on the performance and energy advantages of its custom tensor processor. Read more…

By Tiffany Trader

Nvidia’s Mammoth Volta GPU Aims High for AI, HPC

May 10, 2017

At Nvidia's GPU Technology Conference (GTC17) in San Jose, Calif., this morning, CEO Jensen Huang announced the company's much-anticipated Volta architecture a Read more…

By Tiffany Trader

Facebook Open Sources Caffe2; Nvidia, Intel Rush to Optimize

April 18, 2017

From its F8 developer conference in San Jose, Calif., today, Facebook announced Caffe2, a new open-source, cross-platform framework for deep learning. Caffe2 is the successor to Caffe, the deep learning framework developed by Berkeley AI Research and community contributors. Read more…

By Tiffany Trader

Leading Solution Providers

MIT Mathematician Spins Up 220,000-Core Google Compute Cluster

April 21, 2017

On Thursday, Google announced that MIT math professor and computational number theorist Andrew V. Sutherland had set a record for the largest Google Compute Engine (GCE) job. Sutherland ran the massive mathematics workload on 220,000 GCE cores using preemptible virtual machine instances. Read more…

By Tiffany Trader

Google Debuts TPU v2 and will Add to Google Cloud

May 25, 2017

Not long after stirring attention in the deep learning/AI community by revealing the details of its Tensor Processing Unit (TPU), Google last week announced the Read more…

By John Russell

US Supercomputing Leaders Tackle the China Question

March 15, 2017

Joint DOE-NSA report responds to the increased global pressures impacting the competitiveness of U.S. supercomputing. Read more…

By Tiffany Trader

Groq This: New AI Chips to Give GPUs a Run for Deep Learning Money

April 24, 2017

CPUs and GPUs, move over. Thanks to recent revelations surrounding Google’s new Tensor Processing Unit (TPU), the computing world appears to be on the cusp of Read more…

By Alex Woodie

Russian Researchers Claim First Quantum-Safe Blockchain

May 25, 2017

The Russian Quantum Center today announced it has overcome the threat of quantum cryptography by creating the first quantum-safe blockchain, securing cryptocurrencies like Bitcoin, along with classified government communications and other sensitive digital transfers. Read more…

By Doug Black

DOE Supercomputer Achieves Record 45-Qubit Quantum Simulation

April 13, 2017

In order to simulate larger and larger quantum systems and usher in an age of “quantum supremacy,” researchers are stretching the limits of today’s most advanced supercomputers. Read more…

By Tiffany Trader

Messina Update: The US Path to Exascale in 16 Slides

April 26, 2017

Paul Messina, director of the U.S. Exascale Computing Project, provided a wide-ranging review of ECP’s evolving plans last week at the HPC User Forum. Read more…

By John Russell

Knights Landing Processor with Omni-Path Makes Cloud Debut

April 18, 2017

HPC cloud specialist Rescale is partnering with Intel and HPC resource provider R Systems to offer first-ever cloud access to Xeon Phi "Knights Landing" processors. The infrastructure is based on the 68-core Intel Knights Landing processor with integrated Omni-Path fabric (the 7250F Xeon Phi). Read more…

By Tiffany Trader

  • arrow
  • Click Here for More Headlines
  • arrow
Share This