Green500 Turns Blue

By Michael Feldman

July 5, 2012

The latest Green500 rankings were announced last week, revealing that top performance and power efficiency can indeed go hand in hand. According to the latest list, the greenest machines, in fact the top 20 systems, were all IBM Blue Gene/Q supercomputers. Blue Gene/Q, of course, is the platform that captured the number one spot on the latest TOP500 list, and is represented by four of the ten fastest supercomputers in the world.

Not only did Blue Gene/Q dominate the top of the Green500, but did so in commanding fashion. Although the smaller Q machines tended be slightly more energy efficient, all 20 delivered more than 2,000 megaflops/watt. That turned out to be about twice as efficient as the average for the next 20 supercomputers on the list.

Of those following 20 systems, 10 are accelerator-based. In fact, the 21st and 22nd most efficient supercomputers are the Intel MIC-accelerated prototype cluster (1380.67 megaflops/watt) and the ATI Radeon-equipped DEGIMA cluster (1379.79 megaflops/per watt). The remainder all use NVIDIA GPU parts and are only somewhat less power-efficient.

It’s hard to draw a lot of conclusions about the efficiency of accelerator-equipped machines, since the ratio between the more energy-efficient GPUs (or MIC coprocessors) and the CPUs on these machines has a big impact on the overall results. In other words, a high GPU:CPU ratio system would tend to be yield more megaflops/watt than one with a lower ratio. Further, the current crop of accelerator-based systems tend to yield sub-par Linpack performance (the basis of both the TOP500 and Green500 results) compared to the machine’s peak performance, although this “bias” does point out that it can be difficult to extract performance and performance per watt from these heterogeneous platforms.

A number of x86 CPU-only systems, especially those employing the latest Intel “Sandy Bridge” processors, did rather well the latest rankings. In this category is the new 2.9 petaflop SuperMUC machine that just booted up at Germany’s Leibniz Supercomputing Centre (LRZ) . This IBM iDataPlex cluster sits at number 4 on the TOP500 list and manages a very respectable number 39 placement on the Green500. The system uses an innovative hot-water cooling system that not only saves energy costs, but whose waste heat is repurposed for local use at the LRZ facility. The machine also employs system software that is designed to optimize energy consumption.

The other instructive lesson of SuperMUC is that institutions are willing to endure relatively high energy costs to get leading-edge performance. (SuperMUC is currently the speediest supercomputer in Europe.) Even though its innovative cooling system will supposedly save around a million Euros per year, in energy costs, the high price of electricity in Germany will still make SuperMUC the most expensive supercomputer in Europe to operate.

According to Arndt Bode, LRZ’s chairman of the board who spoke about the new system at ISC’12, energy costs for them are rather steep — 0.158 €/kilowatt-hour as of 2010. That’s around 10 times the cost at Oak Ridge National Laboratory, perhaps the least expensive place to do supercomputing in the US, thanks in large part to cheap blocks of power that can be purchased from the Tennessee Valley Authority. Since SuperMUC consumes 3.4 megawatts, that means the Germans are paying for what an equivalent 34 megawatt system would cost the Oak Ridge boys today.

Considering that supercomputing designers have drawn a 20MW line in the sand for exascale systems, the Germans, in effect, have already resigned themselves to that level of cost. Of course, not everyone is going to be able to plop their exascale systems in the Tennessee Valley (or at other cheap energy locales like Iceland). And energy prices are likely to rise between now and the end of the decade, almost everywhere. But 20MW or more (maybe significantly more) is doable for at least some geographies today, assuming the HPC funding and political will is there.

Anyway you look at it, exaflops-level supercomputing is going to be an expensive proposition, at least initially. The average price of a megawatt in the US is a million dollars per year, and even at Oak Ridge, it probably costs between $200 to $300 thousand. That’s plenty of motivation to reduce the energy footprint of these machines.

Which brings us back to Blue Gene/Q. The largest such system today, the number one Sequoia machine at Lawrence Livermore, delivers 20 (peak) petaflops and draws 7.9MW when it’s running floating-point heavy codes like Linpack. But it would need to be 50 times more powerful to get to an exaflop and would also have to be 25 times as energy-efficient to squeeze such a machine into 20MW.

IBM appears to be on the right track here though, at least from the processor standpoint. Unlike a conventional x86-based HPC cluster, Blue Gene Q is powered by a custom SoC based on the PowerPC A2 core. That processor merges the network and compute on-chip, and is designed as a low-power, high throughput, and high core count (16) architecture. Clock frequency is a modest 1.6 GHz, which is about half that of a top bin Xeon. All exascale processors are likely to follow this general design.

It’s not all up to the processor, however. Memory and system network components will also need analogous redesigns to address their own power issues for exascale. By the way, it would be instructive if the Green500 could expand its mandate and develop useful performance per watt metrics aimed at main memory and interconnects. Linpack is a notoriously bad measurement for data movement, which has become the limiting factor for many applications, “big data” and otherwise. A starting point might be to incorporate the Graph 500 results into a separate set of Green500 rankings.

In the meantime, the list is drawing some much-needed attention to HPC power issues. And competition for those top Green500 spots is going to heat up. In the absence of a Blue Gene/R follow-on — and at this point, IBM has kept mum about extending the BG franchise — there is likely to be some stiff competition from machines powered by the upcoming NVIDIA Kepler K20 GPUs and Intel MIC coprocessors, and their successors. AMD APU-based systems might show up in a couple of years, and the newer SPARC64 offerings from Fujitsu or Chinese systems based on domestically designed chips like Godson may make their presence felt as well. The green revolution in HPC is just beginning.

Related Articles

TOP500 Gets Dressed Up with New Blue Genes

HPC Lists We’d Like to See

IBM Specs Out Blue Gene/Q Chip

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industy updates delivered to you every week!

Cray Completes ClusterStor Deal, Sunsets Sonexion Brand

September 25, 2017

Having today completed the transaction and strategic partnership with Seagate announced back in July, Cray is now home to the ClusterStor line and will be sunsetting the Sonexion brand. This is not an acquisition; the ClusterStor assets are transferring from Seagate to Cray (minus the Seagate ClusterStor IBM Spectrum Scale product) and Cray is taking over support and maintenance for the entire ClusterStor base. Read more…

By Tiffany Trader

China’s TianHe-2A will Use Proprietary Accelerator and Boast 94 Petaflops Peak

September 25, 2017

The details of China’s upgrade to TianHe-2 (MilkyWay-2) – now TianHe-2A – were revealed last week at the Third International High Performance Computing Forum (IHPCF2017) in China. The TianHe-2A will use a proprieta Read more…

By John Russell

SC17 Preview: Invited Talk Lineup Includes Gordon Bell, Paul Messina and Many Others

September 25, 2017

With the addition of esteemed supercomputing pioneer Gordon Bell to its invited talk lineup, SC17 now boasts a total of 12 invited talks on its agenda. As SC explains, "Invited Talks are a premier component of the SC Read more…

By Tiffany Trader

HPE Extreme Performance Solutions

HPE Prepares Customers for Success with the HPC Software Portfolio

High performance computing (HPC) software is key to harnessing the full power of HPC environments. Development and management tools enable IT departments to streamline installation and maintenance of their systems as well as create, optimize, and run their HPC applications. Read more…

GlobalFoundries Puts Wind in AMD’s Sails with 12nm FinFET

September 24, 2017

From its annual tech conference last week (Sept. 20), where GlobalFoundries welcomed more than 600 semiconductor professionals (reaching the Santa Clara venue’s max capacity and doubling 2016 attendee numbers), the one Read more…

By Tiffany Trader

Cray Completes ClusterStor Deal, Sunsets Sonexion Brand

September 25, 2017

Having today completed the transaction and strategic partnership with Seagate announced back in July, Cray is now home to the ClusterStor line and will be sunsetting the Sonexion brand. This is not an acquisition; the ClusterStor assets are transferring from Seagate to Cray (minus the Seagate ClusterStor IBM Spectrum Scale product) and Cray is taking over support and maintenance for the entire ClusterStor base. Read more…

By Tiffany Trader

China’s TianHe-2A will Use Proprietary Accelerator and Boast 94 Petaflops Peak

September 25, 2017

The details of China’s upgrade to TianHe-2 (MilkyWay-2) – now TianHe-2A – were revealed last week at the Third International High Performance Computing Fo Read more…

By John Russell

GlobalFoundries Puts Wind in AMD’s Sails with 12nm FinFET

September 24, 2017

From its annual tech conference last week (Sept. 20), where GlobalFoundries welcomed more than 600 semiconductor professionals (reaching the Santa Clara venue Read more…

By Tiffany Trader

Machine Learning at HPC User Forum: Drilling into Specific Use Cases

September 22, 2017

The 66th HPC User Forum held September 5-7, in Milwaukee, Wisconsin, at the elegant and historic Pfister Hotel, highlighting the 1893 Victorian décor and art o Read more…

By Arno Kolster

Stanford University and UberCloud Achieve Breakthrough in Living Heart Simulations

September 21, 2017

Cardiac arrhythmia can be an undesirable and potentially lethal side effect of drugs. During this condition, the electrical activity of the heart turns chaotic, Read more…

By Wolfgang Gentzsch, UberCloud, and Francisco Sahli, Stanford University

PNNL’s Center for Advanced Tech Evaluation Seeks Wider HPC Community Ties

September 21, 2017

Two years ago the Department of Energy established the Center for Advanced Technology Evaluation (CENATE) at Pacific Northwest National Laboratory (PNNL). CENAT Read more…

By John Russell

Exascale Computing Project Names Doug Kothe as Director

September 20, 2017

The Department of Energy’s Exascale Computing Project (ECP) has named Doug Kothe as its new director effective October 1. He replaces Paul Messina, who is stepping down after two years to return to Argonne National Laboratory. Kothe is a 32-year veteran of DOE’s National Laboratory System. Read more…

Takeaways from the Milwaukee HPC User Forum

September 19, 2017

Milwaukee’s elegant Pfister Hotel hosted approximately 100 attendees for the 66th HPC User Forum (September 5-7, 2017). In the original home city of Pabst Blu Read more…

By Merle Giles

How ‘Knights Mill’ Gets Its Deep Learning Flops

June 22, 2017

Intel, the subject of much speculation regarding the delayed, rewritten or potentially canceled “Aurora” contract (the Argonne Lab part of the CORAL “ Read more…

By Tiffany Trader

Reinders: “AVX-512 May Be a Hidden Gem” in Intel Xeon Scalable Processors

June 29, 2017

Imagine if we could use vector processing on something other than just floating point problems.  Today, GPUs and CPUs work tirelessly to accelerate algorithms Read more…

By James Reinders

NERSC Scales Scientific Deep Learning to 15 Petaflops

August 28, 2017

A collaborative effort between Intel, NERSC and Stanford has delivered the first 15-petaflops deep learning software running on HPC platforms and is, according Read more…

By Rob Farber

Oracle Layoffs Reportedly Hit SPARC and Solaris Hard

September 7, 2017

Oracle’s latest layoffs have many wondering if this is the end of the line for the SPARC processor and Solaris OS development. As reported by multiple sources Read more…

By John Russell

Six Exascale PathForward Vendors Selected; DoE Providing $258M

June 15, 2017

The much-anticipated PathForward awards for hardware R&D in support of the Exascale Computing Project were announced today with six vendors selected – AMD Read more…

By John Russell

Top500 Results: Latest List Trends and What’s in Store

June 19, 2017

Greetings from Frankfurt and the 2017 International Supercomputing Conference where the latest Top500 list has just been revealed. Although there were no major Read more…

By Tiffany Trader

IBM Clears Path to 5nm with Silicon Nanosheets

June 5, 2017

Two years since announcing the industry’s first 7nm node test chip, IBM and its research alliance partners GlobalFoundries and Samsung have developed a proces Read more…

By Tiffany Trader

Nvidia Responds to Google TPU Benchmarking

April 10, 2017

Nvidia highlights strengths of its newest GPU silicon in response to Google's report on the performance and energy advantages of its custom tensor processor. Read more…

By Tiffany Trader

Leading Solution Providers

Graphcore Readies Launch of 16nm Colossus-IPU Chip

July 20, 2017

A second $30 million funding round for U.K. AI chip developer Graphcore sets up the company to go to market with its “intelligent processing unit” (IPU) in Read more…

By Tiffany Trader

Google Releases Deeplearn.js to Further Democratize Machine Learning

August 17, 2017

Spreading the use of machine learning tools is one of the goals of Google’s PAIR (People + AI Research) initiative, which was introduced in early July. Last w Read more…

By John Russell

EU Funds 20 Million Euro ARM+FPGA Exascale Project

September 7, 2017

At the Barcelona Supercomputer Centre on Wednesday (Sept. 6), 16 partners gathered to launch the EuroEXA project, which invests €20 million over three-and-a-half years into exascale-focused research and development. Led by the Horizon 2020 program, EuroEXA picks up the banner of a triad of partner projects — ExaNeSt, EcoScale and ExaNoDe — building on their work... Read more…

By Tiffany Trader

Amazon Debuts New AMD-based GPU Instances for Graphics Acceleration

September 12, 2017

Last week Amazon Web Services (AWS) streaming service, AppStream 2.0, introduced a new GPU instance called Graphics Design intended to accelerate graphics. The Read more…

By John Russell

Cray Moves to Acquire the Seagate ClusterStor Line

July 28, 2017

This week Cray announced that it is picking up Seagate's ClusterStor HPC storage array business for an undisclosed sum. "In short we're effectively transitioning the bulk of the ClusterStor product line to Cray," said CEO Peter Ungaro. Read more…

By Tiffany Trader

Russian Researchers Claim First Quantum-Safe Blockchain

May 25, 2017

The Russian Quantum Center today announced it has overcome the threat of quantum cryptography by creating the first quantum-safe blockchain, securing cryptocurrencies like Bitcoin, along with classified government communications and other sensitive digital transfers. Read more…

By Doug Black

GlobalFoundries: 7nm Chips Coming in 2018, EUV in 2019

June 13, 2017

GlobalFoundries has formally announced that its 7nm technology is ready for customer engagement with product tape outs expected for the first half of 2018. The Read more…

By Tiffany Trader

IBM Advances Web-based Quantum Programming

September 5, 2017

IBM Research is pairing its Jupyter-based Data Science Experience notebook environment with its cloud-based quantum computer, IBM Q, in hopes of encouraging a new class of entrepreneurial user to solve intractable problems that even exceed the capabilities of the best AI systems. Read more…

By Alex Woodie

  • arrow
  • Click Here for More Headlines
  • arrow
Share This