HP Adds New HPC Server with On-Board GPGPU

By Michael Feldman

October 5, 2010

Hewlett Packard has launched a new purpose-built HPC rack server with a formidable GPGPU capability. That product, the ProLiant SL390s G7, provides more raw FLOPS per square inch than any server HP has delivered to date, and is the basis for the 2.4 petaflop TSUBAME 2.0 supercomputer currently being deployed at the Tokyo Institute of Technology.

HP actually announced the two new servers this week. Besides the SL390s G7, the company also introduced the SL170s G6, a no-frills server aimed at hyper-scale computing deployments. Both the SL390s and SL170s plug into HP’s new ProLiant SL6500 Scalable System chassis, a 4U box that accommodates up to 8 half-width servers. The SL6500 is the upgrade from the SL6000 system announced last year.

The SL170s and SL390s come as skinless trays rather than the typical server boxes encased in metal, top and bottom. This design could catch on as more vendors look to minimize extraneous hardware and come up with ever-denser rack configurations. SGI also uses a similar skinless design in their CloudRack trays.

The more general-purpose of the new Proliant servers is the SL170s G6, a dual-socket or single-socket server that incorporates the latest Intel Xeon Westmere (5600 series) processors in a half-width form factor. Ed Turkel, HP’s manager of business development for its HPC group, describes it as their “lean and mean server,” where scalable performance, serviceability, manageability are the driving concerns. As such, it’s aimed mostly at Web 2.0 and service provider environments, but it’s also quite suitable for embarrassingly-parallel HPC applications, such as portfolio risk analysis and BLAST-based bioinformatics. The base price on this model is $1,559.

But for “true” supercomputing applications, the SL390s G7 is the go-to server. Like its sibling, the SL390s comes with Xeon 5600 processors, but the option to pair the CPUs with up to three on-board NVIDIA “Fermi” 20-series GPUs puts a lot more floating point performance into this design. Customers can choose from either the M2050 or M2070 Tesla GPU modules, the only difference being the amount of graphics memory — 3 GB of GDDR5 for the M2050 versus 6 GB for the M2070. Each GPU module is served by its own PCIe Gen2 x16 channel in order to maximize bandwidth to the graphics chips. At the maximum configuration with all three Fermi GPUs and two Westmere CPUs, a single server delivers on the order of 1 teraflop of double precision performance. “So this is very much a server that has been designed for HPC,” said Turkel.

With GPUs on board, the SL390s fill out a 2U half-width tray, so up to four of these can be packed into a 4U SL6500 chassis. A CPU-only version is also available and takes up just half the space (half-width 1U), enabling twice as many Xeons to occupy the same chassis. This configuration will likely be the server of choice for the majority of HPC setups, given that GPGPU deployment is really just getting started. Pricing on the CPU-only model starts at $2,259.

Another HPC-centric feature on the SL390s is the inclusion of on-board network adapters, in this case Mellanox’s ConnectX-2 silicon. The embedded adapter supports either 40 Gbps InfiniBand or 10 Gigabit Ethernet, making it suitable for low-latency applications on either fabric. If dual-rail InfiniBand is desired, an external adapter can be hooked into the server’s PCIe slot. The Mellanox silicon has also been incorporated in HP’s ProLiant BL2x220 G7 server blade.

Although the official debut of the SL390s was on Tuesday, HP has been shipping the server for some time, most notably to Tokyo Institute of Technology (Tokyo Tech), where it serves as the foundation for the 2.4 petaflop TSUBAME 2.0 supercomputer. That system is now fully deployed and will be formally launched later this week.

The new TSUBAME consists of 1,432 SL390s G7 servers, each of which contains three M2050 GPUs. CPU-wise, each server is outfitted with two 6-core Westmere processors (X5670, 2.93 MHz) and either 54 or 96 GB of RAM. For ultra-fast local storage, two SSDs plug into each server node. The network fabric is all QDR InfiniBand, taking advantage of the on-board Mellanox chips; an additional InfiniBand adapter is plugged into each node to provide dual-rail InfiniBand. The whole fabric delivers a system-wide aggregate bandwidth of 200 terabits per second.

The servers are housed in HP’s 42U Module Cooling System G2 rack, which represents the basic building block for TSUBAME’s computing infrastructure. Each rack contains 30 SL390s G7 nodes (60, CPUs and 90 GPUs), 8 chassis of power management, an HP network switch for shared console and the local area network, two airflow dams, and 4 Voltaire 4036 leaf switches.

A single rack consumes around 35 KW, 20 KW of which are from the GPUs alone. Not surprisingly, the G2 rack is water cooled, to handle the considerable heat generated by the CPU-GPU configuration. All this makes for a very computationally dense system, and despite the 35 KW power draw per rack, results in a rather efficient supercomputer for both space and energy consumption. “They wanted a world-class system, but they wanted it to fit into 200 square meters of floor space and into 1.8 MW of power,” explained Turkel.

Besides Tokyo Tech, the SL390s G7 has attracted some other early customers. Although Turkel couldn’t name names, he said the new HPC server is already garnering a lot of interest from scientific research organizations, oil and gas firms, and financial services institutions.

With the new server, HP joins IBM, Dell, SGI and just about every other HPC system vendor with on-board GPGPU. Although HP has offered plug-in Tesla cards for its servers and even qualified NVIDIA’s 1U quad-GPU box in the past, the SL390s G7 represents the company’s first generally-available native GPU server design. According to Turkel, there will be other variations of GPGPU-equipped rack servers in the future, but he was noncommittal regarding any plans to offer this capability in HP’s blade server line. Turkel did say the company is aware that NVIDIA will soon be shipping the compact X2070 Tesla module designed specifically for blades and other small form-factor designs, admitting “we’re certainly looking at that.”

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industy updates delivered to you every week!

Microsoft Wants to Speed Quantum Development

December 12, 2017

Quantum computing continues to make headlines in what remains of 2017 as tech giants jockey to establish a pole position in the race toward commercialization of quantum. This week, Microsoft took the next step in advanci Read more…

By Tiffany Trader

ESnet Now Moving More Than 1 Petabyte/wk

December 12, 2017

Optimizing ESnet (Energy Sciences Network), the world's fastest network for science, is an ongoing process. Recently a two-year collaboration by ESnet users – the Petascale DTN Project – achieved its ambitious goal t Read more…

HPC-as-a-Service Finds Toehold in Iceland

December 11, 2017

While high-demand workloads (e.g., bitcoin mining) can overheat data center cooling capabilities, at least one data center infrastructure provider has announced an HPC-as-a-service offering that features 100 percent fre Read more…

By Doug Black

HPE Extreme Performance Solutions

Explore the Origins of Space with COSMOS and Memory-Driven Computing

From the formation of black holes to the origins of space, data is the key to unlocking the secrets of the early universe. Read more…

HPC Iron, Soft, Data, People – It Takes an Ecosystem!

December 11, 2017

Cutting edge advanced computing hardware (aka big iron) does not stand by itself. These computers are the pinnacle of a myriad of technologies that must be carefully woven together by people to create the computational c Read more…

By Alex R. Larzelere

Microsoft Wants to Speed Quantum Development

December 12, 2017

Quantum computing continues to make headlines in what remains of 2017 as tech giants jockey to establish a pole position in the race toward commercialization of Read more…

By Tiffany Trader

HPC Iron, Soft, Data, People – It Takes an Ecosystem!

December 11, 2017

Cutting edge advanced computing hardware (aka big iron) does not stand by itself. These computers are the pinnacle of a myriad of technologies that must be care Read more…

By Alex R. Larzelere

IBM Begins Power9 Rollout with Backing from DOE, Google

December 6, 2017

After over a year of buildup, IBM is unveiling its first Power9 system based on the same architecture as the Department of Energy CORAL supercomputers, Summit a Read more…

By Tiffany Trader

Microsoft Spins Cycle Computing into Core Azure Product

December 5, 2017

Last August, cloud giant Microsoft acquired HPC cloud orchestration pioneer Cycle Computing. Since then the focus has been on integrating Cycle’s organization Read more…

By John Russell

GlobalFoundries, Ayar Labs Team Up to Commercialize Optical I/O

December 4, 2017

GlobalFoundries (GF) and Ayar Labs, a startup focused on using light, instead of electricity, to transfer data between chips, today announced they've entered in Read more…

By Tiffany Trader

HPE In-Memory Platform Comes to COSMOS

November 30, 2017

Hewlett Packard Enterprise is on a mission to accelerate space research. In August, it sent the first commercial-off-the-shelf HPC system into space for testing Read more…

By Tiffany Trader

SC17 Cluster Competition: Who Won and Why? Results Analyzed and Over-Analyzed

November 28, 2017

Everyone by now knows that Nanyang Technological University of Singapore (NTU) took home the highest LINPACK Award and the Overall Championship from the recently concluded SC17 Student Cluster Competition. We also already know how the teams did in the Highest LINPACK and Highest HPCG competitions, with Nanyang grabbing bragging rights for both benchmarks. Read more…

By Dan Olds

Perspective: What Really Happened at SC17?

November 22, 2017

SC is over. Now comes the myriad of follow-ups. Inboxes are filled with templated emails from vendors and other exhibitors hoping to win a place in the post-SC thinking of booth visitors. Attendees of tutorials, workshops and other technical sessions will be inundated with requests for feedback. Read more…

By Andrew Jones

US Coalesces Plans for First Exascale Supercomputer: Aurora in 2021

September 27, 2017

At the Advanced Scientific Computing Advisory Committee (ASCAC) meeting, in Arlington, Va., yesterday (Sept. 26), it was revealed that the "Aurora" supercompute Read more…

By Tiffany Trader

NERSC Scales Scientific Deep Learning to 15 Petaflops

August 28, 2017

A collaborative effort between Intel, NERSC and Stanford has delivered the first 15-petaflops deep learning software running on HPC platforms and is, according Read more…

By Rob Farber

Oracle Layoffs Reportedly Hit SPARC and Solaris Hard

September 7, 2017

Oracle’s latest layoffs have many wondering if this is the end of the line for the SPARC processor and Solaris OS development. As reported by multiple sources Read more…

By John Russell

AMD Showcases Growing Portfolio of EPYC and Radeon-based Systems at SC17

November 13, 2017

AMD’s charge back into HPC and the datacenter is on full display at SC17. Having launched the EPYC processor line in June along with its MI25 GPU the focus he Read more…

By John Russell

Nvidia Responds to Google TPU Benchmarking

April 10, 2017

Nvidia highlights strengths of its newest GPU silicon in response to Google's report on the performance and energy advantages of its custom tensor processor. Read more…

By Tiffany Trader

Japan Unveils Quantum Neural Network

November 22, 2017

The U.S. and China are leading the race toward productive quantum computing, but it's early enough that ultimate leadership is still something of an open questi Read more…

By Tiffany Trader

GlobalFoundries Puts Wind in AMD’s Sails with 12nm FinFET

September 24, 2017

From its annual tech conference last week (Sept. 20), where GlobalFoundries welcomed more than 600 semiconductor professionals (reaching the Santa Clara venue Read more…

By Tiffany Trader

Google Releases Deeplearn.js to Further Democratize Machine Learning

August 17, 2017

Spreading the use of machine learning tools is one of the goals of Google’s PAIR (People + AI Research) initiative, which was introduced in early July. Last w Read more…

By John Russell

Leading Solution Providers

Amazon Debuts New AMD-based GPU Instances for Graphics Acceleration

September 12, 2017

Last week Amazon Web Services (AWS) streaming service, AppStream 2.0, introduced a new GPU instance called Graphics Design intended to accelerate graphics. The Read more…

By John Russell

Perspective: What Really Happened at SC17?

November 22, 2017

SC is over. Now comes the myriad of follow-ups. Inboxes are filled with templated emails from vendors and other exhibitors hoping to win a place in the post-SC thinking of booth visitors. Attendees of tutorials, workshops and other technical sessions will be inundated with requests for feedback. Read more…

By Andrew Jones

EU Funds 20 Million Euro ARM+FPGA Exascale Project

September 7, 2017

At the Barcelona Supercomputer Centre on Wednesday (Sept. 6), 16 partners gathered to launch the EuroEXA project, which invests €20 million over three-and-a-half years into exascale-focused research and development. Led by the Horizon 2020 program, EuroEXA picks up the banner of a triad of partner projects — ExaNeSt, EcoScale and ExaNoDe — building on their work... Read more…

By Tiffany Trader

Delays, Smoke, Records & Markets – A Candid Conversation with Cray CEO Peter Ungaro

October 5, 2017

Earlier this month, Tom Tabor, publisher of HPCwire and I had a very personal conversation with Cray CEO Peter Ungaro. Cray has been on something of a Cinderell Read more…

By Tiffany Trader & Tom Tabor

Tensors Come of Age: Why the AI Revolution Will Help HPC

November 13, 2017

Thirty years ago, parallel computing was coming of age. A bitter battle began between stalwart vector computing supporters and advocates of various approaches to parallel computing. IBM skeptic Alan Karp, reacting to announcements of nCUBE’s 1024-microprocessor system and Thinking Machines’ 65,536-element array, made a public $100 wager that no one could get a parallel speedup of over 200 on real HPC workloads. Read more…

By John Gustafson & Lenore Mullin

Flipping the Flops and Reading the Top500 Tea Leaves

November 13, 2017

The 50th edition of the Top500 list, the biannual publication of the world’s fastest supercomputers based on public Linpack benchmarking results, was released Read more…

By Tiffany Trader

IBM Begins Power9 Rollout with Backing from DOE, Google

December 6, 2017

After over a year of buildup, IBM is unveiling its first Power9 system based on the same architecture as the Department of Energy CORAL supercomputers, Summit a Read more…

By Tiffany Trader

Intel Launches Software Tools to Ease FPGA Programming

September 5, 2017

Field Programmable Gate Arrays (FPGAs) have a reputation for being difficult to program, requiring expertise in specialty languages, like Verilog or VHDL. Easin Read more…

By Tiffany Trader

  • arrow
  • Click Here for More Headlines
  • arrow
Share This