Beyond Speeds and Feeds

By By Geoffrey James

July 13, 2009

High Performance Computing (HPC) was once limited to a select group of laboratories where scientists or engineers solved complex problems on huge mainframe “supercomputers” that cost millions of dollars to buy and maintain. Today the drop in the price of computer power has combined with new architectures for clustering to bring HPC to a wide range of applications inside a growing number of industries at a reasonable price.

“Computer power is the raw fuel for business innovation,” explains Dr. Jeff Layton, enterprise technologist for HPC at Dell Inc. “Making HPC available to a wider range of customers, and making it more cost-effective, will have a long-term effect, not just on productivity but also on the ability of companies to thrive, not only during difficult economic times but also for many years to come.”

Along with this democratization of HPC has come a growing understanding, among pundits and executives alike, that the traditional way of measuring HPC—the raw performance of a single CPU—seems out of date. As the computer industry leaps to more-complex computing environments, it has become clear that HPC performance must be redefined in order to encapsulate the wider business case, according to Scot Schultz, AMD’s senior strategic alliance manager for HPC.

“What’s important is not how fast the CPU can run a test suite but how effectively it can solve a real-life problem,” Schultz says.

Productivity Now Trumps Raw Performance
More and more analysts, OEMs and IT executives have come to understand that raw performance is less important than how the underlying architecture makes end users more productive. “The performance that’s actually delivered to end users is highly dependent on the chip architecture and how well the software can take advantage of it,” explains Layton.

IT managers who make HPC buying decisions based purely on those obsolete measurements risk getting less bang for their buck, according to John Spooner, an analyst at the market research firm Technology Business Research (TBR). “There are always going to be customers who want all-out performance and don’t care about anything else,” he admits, “but many companies are now embracing the idea that the greatest business value comes not from raw performance but from getting the maximum performance for your overall IT dollar.”

Companies that adopt HPC are typically less interested in “speed and feeds” than in creating a long-term competitive advantage. A case in point is the sport department of Ferrari, one of the first companies to test Microsoft’s Windows HPC Server 2008.

“Ferrari is always looking for the most-advanced technological solutions, and the same goes for software and engineering,” says Piergiorgio Grossi, head of information systems at Ferrari. Like many other companies embracing HPC today, Ferrari is using it widely across the corporation—“for our users, engineers and administrators,” Grossi says.

Companies need to be thinking about productivity as a performance measurement, according to Vince Mendillo, director of marketing for the HPC business group at Microsoft. “HPC is expanding into vertical markets, ranging from engineering to aerospace to energy and many other industries,” he explains. “Ultimately, HPC is about helping customers get the job done.”

Measuring Productivity
HPC has traditionally been measured in terms of the raw computing power of a single core on a single CPU. Using that primitive metric, the battle for “market leadership” has been primarily between the two leading CPU firms: AMD and Intel, according to Rob Enderle of the Enderle Group. “For decades, these two companies have traded positions as the ‘industry leader’ when it comes to raw performance figures,” he says.

It’s a contest that’s likely to continue for the foreseeable future, according to Ken Cayton, research manager for enterprise platforms at the market research firm IDC. “Both companies are constantly moving forward, so one would expect to see the same kind of leapfrog behavior we’ve seen so frequently in the past,” he says.

However, IT executives need to be aware that the traditional “speeds and feeds” measurement is largely irrelevant in a world in which HPC takes place on CPU chips that contain multiple cores, which are, in turn, harnessed into clusters. In a multiprocessing environment, other metrics such as power efficiency start becoming more important, according to TBR’s Spooner. “Because energy costs are such a big proportion of the expense of running a large data center, businesses now want to maximize the amount of work they get done for each unit of electricity they pay for,” he explains.

Indeed, some companies are finding that the hard limitation of their HPC computing isn’t raw performance but the amount of electricity they can get piped into their data center. Cayton relates an experience he recently had with a Manhattan firm that is doing financial analysis but has only a limited amount of power coming into the building. “It therefore is more concerned with how effectively its HPC system uses power than it is about how quickly one element of the system can perform calculations,” Cayton explains.

Architecture and Performance
With multiprocessing and clustering, the speed of an individual core is often far less important than the ability to move data around between the various chips, explains Jordan Selburn, principal analyst at the market research firm iSuppli.

“In a lot of areas and applications, raw horsepower isn’t a significant factor, because other standards drive the degree of speed needed and anything excess is just that: excess,” Selburn explains. “The key in HPC applications is how efficiently you can perform the needed function.”

And that efficiency is intimately tied to the underlying architecture of the CPU chip, according to Einar Rustad, vice president of business development at Numascale, a company that makes chip sets that link multiple CPUs into HPC clusters. “The challenge with multiprocessing is keeping everything in sync, which means that each CPU must have swift access to the data that’s been processed by the other CPUs,” he explains.

To accomplish this, the cluster must be able to move data around quickly, something the HyperTransport™ architecture that AMD uses makes relatively easy. “With other chip architectures, you have to move data around by using the front-side bus, which is not only ungainly from an electronics viewpoint but also incurs a lot of overhead and prevents a true shared memory architecture with cache coherence,” says Rustad. “AMD’s HyperTransport technology, by contrast, makes it easier to connect CPUs together in a way that enables programmers to address the combined memory space and to benefit from the aggregated memory bandwidth.”

One benefit of directly connecting the chips is a potential decrease in data latency, which means that each CPU in the cluster will spend less time idling and more time actually processing data, according to Gilad Shainer, director of technical marketing for Mellanox Technologies, a leading supplier of semiconductor-based server and storage interconnect products.

“AMD has a good vision of how HPC should be handled,” Shainer says. “Its technological architecture provides value for many applications and end users, which is why we’re happy to collaborate with it to build the kind of balanced systems that companies want to buy.”

Real-Life Productivity
A chip architecture that handles data more efficiently can also make life easier for HPC programmers—an important issue in IT groups that may have limited access to top programming talent.

“One of the big limitations in HPC is adapting programs to run in parallel,” says Dell’s Layton. “The computer industry has been struggling for years with limitations on memory bandwidth per core, but that’s finally beginning to ease up, largely as the result of improvements in basic CPU architecture.”

Because programming for HPC is becoming easier, it’s beginning to show up in more industries and application areas. And that, in turn, has further lessened the importance of raw computing power as the primary HPC benchmark, because every industry has different requirements when it comes to the type of computing power that’s applicable to that industry. For example, financial HPC applications make extensive use of floating point, an area in which AMD’s architecture has a “slight edge” over other architectures, according to Christian Heidarson, an analyst at the market research firm Gartner.

HPC-friendly chip architecture can also make a future upgrade path easier. “Because HPC applications tend to be complex, companies are leery of pulling out their current systems and replacing them with new ones,” explains Layton, who notes that AMD has been designing CPU architectures that are socket-compatible, making it possible to upgrade a system without reloading and reconfiguring the software. The only change to the system that’s required is a BIOS upgrade, which takes a few minutes as opposed to the hours or days it might take to completely reconstruct a clustered system. “This makes it possible for a company to upgrade while limiting the downtime and cost risks inherent in re-creating and reinitializing the entire cluster,” says Layton.

In short, the raw performance of a single CPU may not be the best measurement of HPC. Rather, a metric such as the total cost of ownership (TCO) can provide a better baseline by which to judge systems and their underlying chip architecture.

“It’s a big change from the way people are used to thinking about HPC,” says AMD’s Schultz. “However, focusing on productivity means that companies can purchase their computer power more wisely and get the most benefit from their IT dollars.

For more on HPC solutions based on AMD Opteron™ processors go to www.amd.com/istanbulsolutions.

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industy updates delivered to you every week!

Lenovo to Debut ‘Neptune’ Cooling Technologies at ISC ‘18

June 19, 2018

Lenovo today announced a set of cooling technologies, dubbed Neptune, that include direct to node (DTN) warm water cooling, rear door heat exchanger (RDHX), and hybrid solutions that combine air and liquid cooling. Lenov Read more…

By John Russell

World Cup is Lame Compared to This Competition

June 18, 2018

So you think World Cup soccer is a big deal? While I’m sure it’s very compelling to watch a bunch of athletes kick a ball around, World Cup misses the boat because it doesn’t include teams putting together their ow Read more…

By Dan Olds

IBM Demonstrates Deep Neural Network Training with Analog Memory Devices

June 18, 2018

From smarter, more personalized apps to seemingly-ubiquitous Google Assistant and Alexa devices, AI adoption is showing no signs of slowing down – and yet, the hardware used for AI is far from perfect. Currently, GPUs Read more…

By Oliver Peckham

HPE Extreme Performance Solutions

HPC and AI Convergence is Accelerating New Levels of Intelligence

Data analytics is the most valuable tool in the digital marketplace – so much so that organizations are employing high performance computing (HPC) capabilities to rapidly collect, share, and analyze endless streams of data. Read more…

IBM Accelerated Insights

Banks Boost Infrastructure to Tackle GDPR

As banks become more digital and data-driven, their IT managers are challenged with fast growing data volumes and lines-of-businesses’ (LoBs’) seemingly limitless appetite for analytics. Read more…

Sandia to Take Delivery of World’s Largest Arm System

June 18, 2018

While the enterprise remains circumspect on prospects for Arm servers in the datacenter, the leadership HPC community is taking a bolder, brighter view of the x86 server CPU alternative. Amongst current and planned Arm HPC installations – i.e., the innovative Mont-Blanc project, led by Bull/Atos, the 'Isambard’ Cray XC50 going into the University of Bristol, and commitments from both Japan and France among others -- HPE is announcing that it will be supply the United States National Nuclear Security Administration (NNSA) with a 2.3 petaflops peak Arm-based system, named Astra. Read more…

By Tiffany Trader

Sandia to Take Delivery of World’s Largest Arm System

June 18, 2018

While the enterprise remains circumspect on prospects for Arm servers in the datacenter, the leadership HPC community is taking a bolder, brighter view of the x86 server CPU alternative. Amongst current and planned Arm HPC installations – i.e., the innovative Mont-Blanc project, led by Bull/Atos, the 'Isambard’ Cray XC50 going into the University of Bristol, and commitments from both Japan and France among others -- HPE is announcing that it will be supply the United States National Nuclear Security Administration (NNSA) with a 2.3 petaflops peak Arm-based system, named Astra. Read more…

By Tiffany Trader

The Machine Learning Hype Cycle and HPC

June 14, 2018

Like many other HPC professionals I’m following the hype cycle around machine learning/deep learning with interest. I subscribe to the view that we’re probably approaching the ‘peak of inflated expectation’ but not quite yet starting the descent into the ‘trough of disillusionment. This still raises the probability that... Read more…

By Dairsie Latimer

Xiaoxiang Zhu Receives the 2018 PRACE Ada Lovelace Award for HPC

June 13, 2018

Xiaoxiang Zhu, who works for the German Aerospace Center (DLR) and Technical University of Munich (TUM), was awarded the 2018 PRACE Ada Lovelace Award for HPC for her outstanding contributions in the field of high performance computing (HPC) in Europe. Read more…

By Elizabeth Leake

U.S Considering Launch of National Quantum Initiative

June 11, 2018

Sometime this month the U.S. House Science Committee will introduce legislation to launch a 10-year National Quantum Initiative, according to a recent report by Read more…

By John Russell

ORNL Summit Supercomputer Is Officially Here

June 8, 2018

Oak Ridge National Laboratory (ORNL) together with IBM and Nvidia celebrated the official unveiling of the Department of Energy (DOE) Summit supercomputer toda Read more…

By Tiffany Trader

Exascale USA – Continuing to Move Forward

June 6, 2018

The end of May 2018, saw several important events that continue to advance the Department of Energy’s (DOE) Exascale Computing Initiative (ECI) for the United Read more…

By Alex R. Larzelere

Exascale for the Rest of Us: Exaflops Systems Capable for Industry

June 6, 2018

Enterprise advanced scale computing – or HPC in the enterprise – is an entity unto itself, situated between (and with characteristics of) conventional enter Read more…

By Doug Black

Fracas in Frankfurt: ISC18 Cluster Competition Teams Unveiled

June 6, 2018

The Student Cluster Competition season heats up with the seventh edition of the ISC Student Cluster Competition, slated to begin on June 25th in Frankfurt, Germ Read more…

By Dan Olds

MLPerf – Will New Machine Learning Benchmark Help Propel AI Forward?

May 2, 2018

Let the AI benchmarking wars begin. Today, a diverse group from academia and industry – Google, Baidu, Intel, AMD, Harvard, and Stanford among them – releas Read more…

By John Russell

How the Cloud Is Falling Short for HPC

March 15, 2018

The last couple of years have seen cloud computing gradually build some legitimacy within the HPC world, but still the HPC industry lies far behind enterprise I Read more…

By Chris Downing

US Plans $1.8 Billion Spend on DOE Exascale Supercomputing

April 11, 2018

On Monday, the United States Department of Energy announced its intention to procure up to three exascale supercomputers at a cost of up to $1.8 billion with th Read more…

By Tiffany Trader

Deep Learning at 15 PFlops Enables Training for Extreme Weather Identification at Scale

March 19, 2018

Petaflop per second deep learning training performance on the NERSC (National Energy Research Scientific Computing Center) Cori supercomputer has given climate Read more…

By Rob Farber

Lenovo Unveils Warm Water Cooled ThinkSystem SD650 in Rampup to LRZ Install

February 22, 2018

This week Lenovo took the wraps off the ThinkSystem SD650 high-density server with third-generation direct water cooling technology developed in tandem with par Read more…

By Tiffany Trader

ORNL Summit Supercomputer Is Officially Here

June 8, 2018

Oak Ridge National Laboratory (ORNL) together with IBM and Nvidia celebrated the official unveiling of the Department of Energy (DOE) Summit supercomputer toda Read more…

By Tiffany Trader

Nvidia Responds to Google TPU Benchmarking

April 10, 2017

Nvidia highlights strengths of its newest GPU silicon in response to Google's report on the performance and energy advantages of its custom tensor processor. Read more…

By Tiffany Trader

HPE Wins $57 Million DoD Supercomputing Contract

February 20, 2018

Hewlett Packard Enterprise (HPE) today revealed details of its massive $57 million HPC contract with the U.S. Department of Defense (DoD). The deal calls for HP Read more…

By Tiffany Trader

Leading Solution Providers

SC17 Booth Video Tours Playlist

Altair @ SC17

Altair

AMD @ SC17

AMD

ASRock Rack @ SC17

ASRock Rack

CEJN @ SC17

CEJN

DDN Storage @ SC17

DDN Storage

Huawei @ SC17

Huawei

IBM @ SC17

IBM

IBM Power Systems @ SC17

IBM Power Systems

Intel @ SC17

Intel

Lenovo @ SC17

Lenovo

Mellanox Technologies @ SC17

Mellanox Technologies

Microsoft @ SC17

Microsoft

Penguin Computing @ SC17

Penguin Computing

Pure Storage @ SC17

Pure Storage

Supericro @ SC17

Supericro

Tyan @ SC17

Tyan

Univa @ SC17

Univa

Hennessy & Patterson: A New Golden Age for Computer Architecture

April 17, 2018

On Monday June 4, 2018, 2017 A.M. Turing Award Winners John L. Hennessy and David A. Patterson will deliver the Turing Lecture at the 45th International Sympo Read more…

By Staff

Google Chases Quantum Supremacy with 72-Qubit Processor

March 7, 2018

Google pulled ahead of the pack this week in the race toward "quantum supremacy," with the introduction of a new 72-qubit quantum processor called Bristlecone. Read more…

By Tiffany Trader

Google I/O 2018: AI Everywhere; TPU 3.0 Delivers 100+ Petaflops but Requires Liquid Cooling

May 9, 2018

All things AI dominated discussion at yesterday’s opening of Google’s I/O 2018 developers meeting covering much of Google's near-term product roadmap. The e Read more…

By John Russell

Nvidia Ups Hardware Game with 16-GPU DGX-2 Server and 18-Port NVSwitch

March 27, 2018

Nvidia unveiled a raft of new products from its annual technology conference in San Jose today, and despite not offering up a new chip architecture, there were still a few surprises in store for HPC hardware aficionados. Read more…

By Tiffany Trader

Pattern Computer – Startup Claims Breakthrough in ‘Pattern Discovery’ Technology

May 23, 2018

If it weren’t for the heavy-hitter technology team behind start-up Pattern Computer, which emerged from stealth today in a live-streamed event from San Franci Read more…

By John Russell

Part One: Deep Dive into 2018 Trends in Life Sciences HPC

March 1, 2018

Life sciences is an interesting lens through which to see HPC. It is perhaps not an obvious choice, given life sciences’ relative newness as a heavy user of H Read more…

By John Russell

Intel Pledges First Commercial Nervana Product ‘Spring Crest’ in 2019

May 24, 2018

At its AI developer conference in San Francisco yesterday, Intel embraced a holistic approach to AI and showed off a broad AI portfolio that includes Xeon processors, Movidius technologies, FPGAs and Intel’s Nervana Neural Network Processors (NNPs), based on the technology it acquired in 2016. Read more…

By Tiffany Trader

Google Charts Two-Dimensional Quantum Course

April 26, 2018

Quantum error correction, essential for achieving universal fault-tolerant quantum computation, is one of the main challenges of the quantum computing field and it’s top of mind for Google’s John Martinis. At a presentation last week at the HPC User Forum in Tucson, Martinis, one of the world's foremost experts in quantum computing, emphasized... Read more…

By Tiffany Trader

  • arrow
  • Click Here for More Headlines
  • arrow
Do NOT follow this link or you will be banned from the site!
Share This