Flipping the Flops and Reading the Top500 Tea Leaves

By Tiffany Trader

November 13, 2017

The 50th edition of the Top500 list, the biannual publication of the world’s fastest supercomputers based on public Linpack benchmarking results, was released from SC17 in Denver, Colorado, this morning and once again China is in the spotlight, having taken what is on the surface at least a definitive lead in multiple dimensions. China now claims the most systems, biggest flops share and the number one machine for 10 consecutive lists. It’s a coup-level achievement to pull off in five years, disrupting 20 years of US dominance on the Top500, but reading deeper into the Top500 tea leaves reveals a more nuanced analysis that has as much to do with China’s benchmarking chops as it does its supercomputing flops.

PEZY-SC2 chip at ISC 2017 –click to enlarge

Before we thread that needle, let’s take a moment to review the movement at the top of the list. There are no new list entrants in the top ten and no change in the top three, but the upgraded ZettaScaler-2.2 “Gyoukou” stuck its landing for a fourth place ranking. Vaulting 65 spots, the supersized Gyoukou combines Xeons and PEZY-SC2 accelerators to achieve 19.14 petaflops, up from 1.68 petaflops on the previous list. The Top500 authors point out that the system’s 19,860,000 cores represent the highest level of concurrency ever recorded on the Top500 rankings.

Gyoukou also had the honor of being the fifth greenest supercomputer. Fellow ZettaScaler systems Shoubu system B, Suiren2 and Sakura, placed first, second and third respectively (see perf-per-watt numbers below). Nvidia’s DGX SaturnV Volta system, installed at Nvidia headquarters in San Jose, Calif., was the fourth greenest supercomputer.

Nov. 2017 Green500 top five — click to enlarge
Nov. 2017 Top500 top 10

Another upgraded machine, Trinity, moved up three positions to seventh place thanks to a recent infusion of Intel Knights Landing Xeon Phi processors that raised its Linpack score from 8.10 petaflops to 14.14 petaflops. Trinity is a Cray XC40 supercomputer operated by Los Alamos National Laboratory and Sandia National Laboratories.

Sunway TaihuLight datacenter (Wuxi, China)

China still has a firm grip on the top of the list with 93-petaflops Sunway TaihuLight and 33.86-petaflops Tianhe-2, the number one and and two systems respectively, which together provide the new list with 15 percent of its flops. Piz Daint, the Cray XC50 system installed at the Swiss National Supercomputing Centre (CSCS) remains the third fastest system with 19.6 petaflops. With Gyoukou in fourth position, the fastest US system, Titan, slips another notch to fifth place, leaving the United States without a claim to any of the top four rankings. Benchmarked at 17.59 petaflops, the five-year-old Cray XK7 system installed at the Department of Energy’s Oak Ridge National Laboratory, captured the top spot for one list iteration before being knocked off its perch in June 2013 by China’s Tianhe-2. This is the first time in the list’s 24-year history that the US has not held at least a number four ranking.

Although China has enjoyed number one bragging rights for nearly four years, this is the first list that it also dominates by both system number and aggregate performance share as well. China has the most installed systems: 202 compared to 159 on the last list, while US is in second place with 144 down from 169 six month ago (Japan ranks third place with 35, followed by Germany with 20, France with 18, and the UK with 15.). Aggregate performance is similar: China holds 35.3 percent of list flops, and the US is second with 29.8 percent (then Japan with 10.8 percent, Germany with 4.5 percent, UK with 3.8 percent and France with 3.6 percent).

Based on these metrics, undoubtedly some publications will proclaim China’s supercomputing supremacy, but that would be premature. When China expanded its Top500 toehold by a factor of three at SC15, Intersect360 Research CEO Addison Snell remarked that it wasn’t so much that China discovered supercomputing as it discovered the Top500 list. This observation continues to hold water.

An examination of the new systems China is adding to the list indicates concerted efforts by Chinese vendors Inspur, Lenovo, Sugon and more recently Huawei to benchmark loosely coupled Web/cloud systems that strain the definition of HPC. To wit, 68 out of the 96 systems that China introduced onto the latest list utilize 10G networking and none are deployed at research sites. The benchmarking of Internet and telecom systems for Top500 glory is not new. You can see similar fingerprints on the list (current and historical) from HPE and IBM, but China has doubled down. For comparison’s sake, the US put 19 new systems on the list and eight of those rely on 10G networking.

Top500 development over time–countries by performance share. US is red; China is dark blue. Click to enlarge.

Not only has the Linpacking of non-HPC systems inflated China’s list presence, it’s changed the networking demographics as the number of Ethernet-based machines climbs steadily. As the Top500 authors note, Gigabit Ethernet now connects 228 systems with 204 systems using 10G interfaces. InfiniBand technology is now found on 163 systems, down from 178 systems six months ago, and is the second most-used internal system interconnect technology.

Snell provided additional perspective: “What we’re seeing is a concerted effort to list systems in China, particularly from China-based system vendors. The submission rules allow for what is essentially benchmarking by proxy. If Linpack is run and verified on one system, the result can be assumed for other systems of the same (or greater) configuration, so it’s possible to put together concerted efforts to list more systems, whether out of a desire to show apparent market share, or simply for national pride.”

Discussions of list purity and benchmarking by proxy aside, the High Performance Linpack or any one-dimensional metric has limited usefulness across today’s broad mix of HPC applications. This truth, well understood in HPC circles, is not always appreciated outside the community or among government stakeholders who want “something to show” for public investment.

“Actual system effectiveness is getting more difficult to compare, as the industry swings back toward specialized hardware,” Snell commented. “Just because one architecture outperforms another on one benchmark doesn’t make it the best choice for all workloads. This is particularly challenging for mixed-workload research environments trying to serve multiple domains. 88 percent of all HPC users say they will need to support multiple architectures for the next few years, running applications on the most appropriate systems for their requirements.”

Chip technology – click to expand (Source: Top500)

There has been stagnation on the list for several iterations and turnover is historically low. Neither Summit or Sierra (the US CORAL machines, projected to achieve 150-200 petaflops) nor the upgraded Tianhe-2A (projected 94.97 petaflops peak) made the cut for the 50th list as had been speculated. While HPC is seeing a time of increased architectural diversity at the system and processor level, the current list is less diverse by some measures. To wit, of the 136 new systems on the list, Intel is foundational to all of them (36 of these utilize accelerators*). So no new Power, no new AMD (it’s still early for EPYC) and nothing from ARM yet. In total 471 systems, or 94.2 percent, are now using Intel processors, up a notch from 92.8 percent six months ago. The share of IBM Power processors is at 14 systems, down from 21 systems in June. There are five AMD-based systems remaining on the list, down from seven one year ago.

Nvidia’s New SaturnV Volta system. Click to enlarge.

In the US, IBM Power9 systems Summit and Sierra are on track for 2018 installation at Oak Ridge and Livermore labs (respectively), and multiple other exascale-focused systems are in play in China, Europe and Japan, showcasing a new wave of architectural diversity. We expect there will be more exciting supercomputing trends to report on from ISC 2017 in Frankfurt.

*Breakdown of the 36 new accelerated systems: 29 have P100s (one with NVLink, an HPE SGI system at number 292 (Japan)), one internal Nvidia V100 Volta system (#149, SaturnV Volta); one K80-based system (#267, Lenovo); two Sugon-built P40 systems (#161, #300), and three PEZY systems (#260, #277, #308). Further, out of the 36, only the internal Nvidia machine is US-based. 30 are Chinese (by Lenovo, Inspur, Sugon); the remaining five are Japanese (by NTT, HPE, PEZY).

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industy updates delivered to you every week!

Royalty-free stock illustration ID: 1675260034

Solving Heterogeneous Programming Challenges with SYCL

December 8, 2021

In the first of a series of guest posts on heterogenous computing, James Reinders, who returned to Intel last year after a short "retirement," considers how SYCL will contribute to a heterogeneous future for C++. Reinde Read more…

Quantinuum Debuts Quantum-based Cryptographic Key Service – Is this Quantum Advantage?

December 7, 2021

Quantinuum – the newly-named company resulting from the merger of Honeywell’s quantum computing division and UK-based Cambridge Quantum – today launched Quantum Origin, a service to deliver “completely unpredicta Read more…

SC21 Was Unlike Any Other — Was That a Good Thing?

December 3, 2021

For a long time, the promised in-person SC21 seemed like an impossible fever dream, the assurances of a prominent physical component persisting across years of canceled conferences, including two virtual ISCs and the virtual SC20. With the advent of the Delta variant, Covid surges in St. Louis and contention over vaccine requirements... Read more…

The Green500’s Crystal Anniversary Sees MN-3 Crystallize Its Winning Streak

December 2, 2021

“This is the 30th Green500,” said Wu Feng, custodian of the Green500 list, at the list’s SC21 birds-of-a-feather session. “You could say 15 years of Green500, which makes it, I guess, the crystal anniversary.” Indeed, HPCwire marked the 15th anniversary of the Green500 – which ranks supercomputers by flops-per-watt, rather than just by flops – earlier this year with... Read more…

AWS Arm-based Graviton3 Instances Now in Preview

December 1, 2021

Three years after unveiling the first generation of its AWS Graviton chip-powered instances in 2018, Amazon Web Services announced that the third generation of the processors – the AWS Graviton3 – will power all-new Amazon Elastic Compute 2 (EC2) C7g instances that are now available in preview. Debuting at the AWS re:Invent 2021... Read more…

AWS Solution Channel

Introducing AWS HPC Connector for NICE EnginFrame

HPC customers regularly tell us about their excitement when they’re starting to use the cloud for the first time. In conversations, we always want to dig a bit deeper to find out how we can improve those initial experiences and deliver on the potential they see. Read more…

Nvidia Dominates Latest MLPerf Results but Competitors Start Speaking Up

December 1, 2021

MLCommons today released its fifth round of MLPerf training benchmark results with Nvidia GPUs again dominating. That said, a few other AI accelerator companies participated and, one of them, Graphcore, even held a separ Read more…

Royalty-free stock illustration ID: 1675260034

Solving Heterogeneous Programming Challenges with SYCL

December 8, 2021

In the first of a series of guest posts on heterogenous computing, James Reinders, who returned to Intel last year after a short "retirement," considers how SYC Read more…

Quantinuum Debuts Quantum-based Cryptographic Key Service – Is this Quantum Advantage?

December 7, 2021

Quantinuum – the newly-named company resulting from the merger of Honeywell’s quantum computing division and UK-based Cambridge Quantum – today launched Q Read more…

SC21 Was Unlike Any Other — Was That a Good Thing?

December 3, 2021

For a long time, the promised in-person SC21 seemed like an impossible fever dream, the assurances of a prominent physical component persisting across years of canceled conferences, including two virtual ISCs and the virtual SC20. With the advent of the Delta variant, Covid surges in St. Louis and contention over vaccine requirements... Read more…

The Green500’s Crystal Anniversary Sees MN-3 Crystallize Its Winning Streak

December 2, 2021

“This is the 30th Green500,” said Wu Feng, custodian of the Green500 list, at the list’s SC21 birds-of-a-feather session. “You could say 15 years of Green500, which makes it, I guess, the crystal anniversary.” Indeed, HPCwire marked the 15th anniversary of the Green500 – which ranks supercomputers by flops-per-watt, rather than just by flops – earlier this year with... Read more…

Nvidia Dominates Latest MLPerf Results but Competitors Start Speaking Up

December 1, 2021

MLCommons today released its fifth round of MLPerf training benchmark results with Nvidia GPUs again dominating. That said, a few other AI accelerator companies Read more…

At SC21, Experts Ask: Can Fast HPC Be Green?

November 30, 2021

HPC is entering a new era: exascale is (somewhat) officially here, but Moore’s law is ending. Power consumption and other sustainability concerns loom over the enormous systems and chips of this new epoch, for both cost and compliance reasons. Reconciling the need to continue the supercomputer scale-up while reducing HPC’s environmental impacts... Read more…

Raja Koduri and Satoshi Matsuoka Discuss the Future of HPC at SC21

November 29, 2021

HPCwire's Managing Editor sits down with Intel's Raja Koduri and Riken's Satoshi Matsuoka in St. Louis for an off-the-cuff conversation about their SC21 experience, what comes after exascale and why they are collaborating. Koduri, senior vice president and general manager of Intel's accelerated computing systems and graphics (AXG) group, leads the team... Read more…

Jack Dongarra on SC21, the Top500 and His Retirement Plans

November 29, 2021

HPCwire's Managing Editor sits down with Jack Dongarra, Top500 co-founder and Distinguished Professor at the University of Tennessee, during SC21 in St. Louis to discuss the 2021 Top500 list, the outlook for global exascale computing, and what exactly is going on in that Viking helmet photo. Read more…

IonQ Is First Quantum Startup to Go Public; Will It be First to Deliver Profits?

November 3, 2021

On October 1 of this year, IonQ became the first pure-play quantum computing start-up to go public. At this writing, the stock (NYSE: IONQ) was around $15 and its market capitalization was roughly $2.89 billion. Co-founder and chief scientist Chris Monroe says it was fun to have a few of the company’s roughly 100 employees travel to New York to ring the opening bell of the New York Stock... Read more…

Enter Dojo: Tesla Reveals Design for Modular Supercomputer & D1 Chip

August 20, 2021

Two months ago, Tesla revealed a massive GPU cluster that it said was “roughly the number five supercomputer in the world,” and which was just a precursor to Tesla’s real supercomputing moonshot: the long-rumored, little-detailed Dojo system. Read more…

Esperanto, Silicon in Hand, Champions the Efficiency of Its 1,092-Core RISC-V Chip

August 27, 2021

Esperanto Technologies made waves last December when it announced ET-SoC-1, a new RISC-V-based chip aimed at machine learning that packed nearly 1,100 cores onto a package small enough to fit six times over on a single PCIe card. Now, Esperanto is back, silicon in-hand and taking aim... Read more…

US Closes in on Exascale: Frontier Installation Is Underway

September 29, 2021

At the Advanced Scientific Computing Advisory Committee (ASCAC) meeting, held by Zoom this week (Sept. 29-30), it was revealed that the Frontier supercomputer is currently being installed at Oak Ridge National Laboratory in Oak Ridge, Tenn. The staff at the Oak Ridge Leadership... Read more…

AMD Launches Milan-X CPU with 3D V-Cache and Multichip Instinct MI200 GPU

November 8, 2021

At a virtual event this morning, AMD CEO Lisa Su unveiled the company’s latest and much-anticipated server products: the new Milan-X CPU, which leverages AMD’s new 3D V-Cache technology; and its new Instinct MI200 GPU, which provides up to 220 compute units across two Infinity Fabric-connected dies, delivering an astounding 47.9 peak double-precision teraflops. “We're in a high-performance computing megacycle, driven by the growing need to deploy additional compute performance... Read more…

Intel Reorgs HPC Group, Creates Two ‘Super Compute’ Groups

October 15, 2021

Following on changes made in June that moved Intel’s HPC unit out of the Data Platform Group and into the newly created Accelerated Computing Systems and Graphics (AXG) business unit, led by Raja Koduri, Intel is making further updates to the HPC group and announcing... Read more…

Killer Instinct: AMD’s Multi-Chip MI200 GPU Readies for a Major Global Debut

October 21, 2021

AMD’s next-generation supercomputer GPU is on its way – and by all appearances, it’s about to make a name for itself. The AMD Radeon Instinct MI200 GPU (a successor to the MI100) will, over the next year, begin to power three massive systems on three continents: the United States’ exascale Frontier system; the European Union’s pre-exascale LUMI system; and Australia’s petascale Setonix system. Read more…

Hot Chips: Here Come the DPUs and IPUs from Arm, Nvidia and Intel

August 25, 2021

The emergence of data processing units (DPU) and infrastructure processing units (IPU) as potentially important pieces in cloud and datacenter architectures was Read more…

Leading Solution Providers

Contributors

D-Wave Embraces Gate-Based Quantum Computing; Charts Path Forward

October 21, 2021

Earlier this month D-Wave Systems, the quantum computing pioneer that has long championed quantum annealing-based quantum computing (and sometimes taken heat fo Read more…

HPE Wins $2B GreenLake HPC-as-a-Service Deal with NSA

September 1, 2021

In the heated, oft-contentious, government IT space, HPE has won a massive $2 billion contract to provide HPC and AI services to the United States’ National Security Agency (NSA). Following on the heels of the now-canceled $10 billion JEDI contract (reissued as JWCC) and a $10 billion... Read more…

The Latest MLPerf Inference Results: Nvidia GPUs Hold Sway but Here Come CPUs and Intel

September 22, 2021

The latest round of MLPerf inference benchmark (v 1.1) results was released today and Nvidia again dominated, sweeping the top spots in the closed (apples-to-ap Read more…

Three Chinese Exascale Systems Detailed at SC21: Two Operational and One Delayed

November 24, 2021

Details about two previously rumored Chinese exascale systems came to light during last week’s SC21 proceedings. Asked about these systems during the Top500 media briefing on Monday, Nov. 15, list author and co-founder Jack Dongarra indicated he was aware of some very impressive results, but withheld comment when asked directly if he had... Read more…

Ahead of ‘Dojo,’ Tesla Reveals Its Massive Precursor Supercomputer

June 22, 2021

In spring 2019, Tesla made cryptic reference to a project called Dojo, a “super-powerful training computer” for video data processing. Then, in summer 2020, Tesla CEO Elon Musk tweeted: “Tesla is developing a [neural network] training computer... Read more…

2021 Gordon Bell Prize Goes to Exascale-Powered Quantum Supremacy Challenge

November 18, 2021

Today at the hybrid virtual/in-person SC21 conference, the organizers announced the winners of the 2021 ACM Gordon Bell Prize: a team of Chinese researchers leveraging the new exascale Sunway system to simulate quantum circuits. The Gordon Bell Prize, which comes with an award of $10,000 courtesy of HPC pioneer Gordon Bell, is awarded annually... Read more…

Quantum Computer Market Headed to $830M in 2024

September 13, 2021

What is one to make of the quantum computing market? Energized (lots of funding) but still chaotic and advancing in unpredictable ways (e.g. competing qubit tec Read more…

IBM Introduces its First Power10-based Server, the Power E1080; Targets Hybrid Cloud

September 8, 2021

IBM today introduced the Power E1080 server, its first system powered by a Power10 IBM microprocessor. The new system reinforces IBM’s emphasis on hybrid cloud markets and the new chip beefs up its inference capabilities. IBM – like other CPU makers – is hoping to make inferencing a core capability... Read more…

  • arrow
  • Click Here for More Headlines
  • arrow
HPCwire