SGI Looks to Freeze Out HPC Competition with New ICE Machine

By Michael Feldman

May 5, 2010

SGI has upgraded its HPC blade server lineup with the latest x86 silicon and a turbo-charged InfiniBand network. The Altix ICE 8400 is the successor to the company’s 8200 series and is designed as a premier solution for the HPC cluster market, scaling as high as 64,000 nodes.

The 8400 represents the fourth generation of the ICE product line, which was begun in 2007, although this iteration is probably the most significant upgrade to the system in its three-year run. Besides moving up to the latest Intel Westmere EP chips (Xeon 5600), for the first time SGI is adding an AMD Opteron option as well. Prior to the company’s merge with Rackable last year, SGI was an Intel-only vendor.

Both x86 blades come in a dual-socket setup and use standard chipsets. The Intel blades feature quad- or six-core Xeon 5600 processors, although customers can opt for the previous generation Xeon 5500 silicon as well. Using the six-core Xeon, up to 768 cores can be stuffed into one cabinet.

The most notable feature on the Xeon side is that SGI has designed the blades to handle the fastest and hottest SKUs in Intel’s arsenal, in other words, the 130 watt parts. To date, all other HPC server offerings have shied away from these top-end Xeon chips since the extra heat produced limits the density of the designs, especially for the closely-packed blade architectures. The standard high-end processors for x86 blades are the 95 watt parts.

“We realize there are some competitors with higher density,” admits Paul Kinyon, senior product marketing manager at SGI. “But what we’ve duly noted is that there is no free lunch.” From their perspective, the HPC market is very sensitive to application licensing costs, and just increasing the CPU server density at the expense of clock speed can end up costing more in software than might have been saved in hardware consolidation.

Speaking of density, the AMD option for the 8400 supports the new 8- and 12-core Opteron 6100 chips (Magny-Cours), which makes a 1,536-core cabinet doable. And since the new Opteron blade supports up to 16 DIMM slots (as opposed to 12 DIMMs on the Xeon blade), there’s more memory to go around as well.

Interestingly, SGI allows you to mix Xeon and Opteron blades in the same cabinet, and run them under the same system manager. A more likely configuration would be to keep the Xeons and Opterons confined to separate racks, using a job scheduler to push specific apps onto the different blades. The rationale is that Intel chips are more suitable to codes needing fewer faster cores, with the AMD chips offering the advantage in memory bandwidth and core count. According to Kinyon, they’ve seen “a fair amount of interest” from customers who are considering a mixed-vendor x86 cluster.

Customers who want to give this x86 odd couple scenario a whirl will have to wait until later in the year, though. While the Intel blades are available now, the AMD hardware won’t be shipping until Q3. From a pure blade perspective (sans CPUs), Kinyon says the AMD and Intel models are similarly priced. Once you add in the CPUs and memory, prices will almost certainly vary. While the Opteron CPUs tend to be less expensive than their Xeon counterparts, if additional memory is desired to support the extra Opteron cores, costs may even out.

Specialized service nodes, which appear as peers to the x86 nodes, can also be integrated into an 8400 cluster. These include shared memory UV10 and NVIDIA Tesla GPU nodes. The shared memory node option, in particular, seems to be gaining traction as an add-on for HPC distributed memory machines, and Kinyon says they’ve already bid this configuration on some recent RFPs.

CPUs and GPUs aside, the bigger story for the new 8400 is what SGI has done with the interconnect. Here they’ve decided to push InfiniBand about as far as it will go. The 8400EX version, in particular, is optimized for maximum interconnect performance. It uses a dual plane network and four integrated QDR InfiniBand switches per enclosure. For better price-performance, the 8400LX offers a single plane network and cuts the InfiniBand switches to two per enclosure.

SGI touts the 8400EX as tops in the industry for MPI performance, delivering a three-fold increase in bandwidth per node versus the competition. The company is claiming a world record result (51.3) for the 8400 on the SPECmpiL_2007 benchmark. Although more pricey, the dual plane design gives customers the option to either use the extra bandwidth as a single fat pipe, employ one of the planes for redundancy for MPI traffic, or dedicate one plane to MPI traffic and the other to I/O.

Multiple network topologies are offered, including hypercube, enhanced hypercube, all-to-all and fat tree topology. Except for the for all-to-all, the other topologies were available in the 8200 product, allowing customers to easily add to their legacy ICE systems by extending the same fabric.

SGI designed the new all-to-all topology to deliver maximum bandwidth (up to 12,000 MB/sec per node) and lowest latency, although this option only scales to 128 nodes. The enhanced hypercube — available in the 8200, but juiced up for the 8400 — is next in bandwidth performance and scales all the way up to tens of thousands of nodes. The fat tree topology is the highest in cost, requiring external switches, but enables all-to-all MPI communication at scale. The hypercube is the lowest cost, but the least performant of the bunch.

According to SGI, there are already some orders on the books for the new Altix. One early deployment is at NASA Ames, where the Pleiades supercomputer just added 32 racks of 8400 hardware, boosting its performance to just shy of a petaflop (973 teraflops).

As seems to be the current tradition in selling high-end servers, SGI is not talking pricing on the 8400. Kinyon says they were very careful when designing the new Altix to make sure that they didn’t price themselves out of the value end of the x86 cluster-based market. At the same time, they wanted to offer a solution for users “on the hairy edge of HPC.” The company believes they’ve struck the right balance of price and performance with the 8400. Says Kinyon: “We’re just jumping up and down and waiting to hear the competition whimper.”

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industy updates delivered to you every week!

Researchers Scale COSMO Climate Code to 4888 GPUs on Piz Daint

October 17, 2017

Effective global climate simulation, sorely needed to anticipate and cope with global warming, has long been computationally challenging. Two of the major obstacles are the needed resolution and prolonged time to compute Read more…

By John Russell

Student Cluster Competition Coverage New Home

October 16, 2017

Hello computer sports fans! This is the first of many (many!) articles covering the world-wide phenomenon of Student Cluster Competitions. Finally, the Student Cluster Competition coverage has come to its natural home: H Read more…

By Dan Olds

UCSD Web-based Tool Tracking CA Wildfires Generates 1.5M Views

October 16, 2017

Tracking the wildfires raging in northern CA is an unpleasant but necessary part of guiding efforts to fight the fires and safely evacuate affected residents. One such tool – Firemap – is a web-based tool developed b Read more…

By John Russell

HPE Extreme Performance Solutions

Transforming Genomic Analytics with HPC-Accelerated Insights

Advancements in the field of genomics are revolutionizing our understanding of human biology, rapidly accelerating the discovery and treatment of genetic diseases, and dramatically improving human health. Read more…

Exascale Imperative: New Movie from HPE Makes a Compelling Case

October 13, 2017

Why is pursuing exascale computing so important? In a new video – Hewlett Packard Enterprise: Eighteen Zeros – four HPE executives, a prominent national lab HPC researcher, and HPCwire managing editor Tiffany Trader Read more…

By John Russell

Student Cluster Competition Coverage New Home

October 16, 2017

Hello computer sports fans! This is the first of many (many!) articles covering the world-wide phenomenon of Student Cluster Competitions. Finally, the Student Read more…

By Dan Olds

Intel Delivers 17-Qubit Quantum Chip to European Research Partner

October 10, 2017

On Tuesday, Intel delivered a 17-qubit superconducting test chip to research partner QuTech, the quantum research institute of Delft University of Technology (TU Delft) in the Netherlands. The announcement marks a major milestone in the 10-year, $50-million collaborative relationship with TU Delft and TNO, the Dutch Organization for Applied Research, to accelerate advancements in quantum computing. Read more…

By Tiffany Trader

Fujitsu Tapped to Build 37-Petaflops ABCI System for AIST

October 10, 2017

Fujitsu announced today it will build the long-planned AI Bridging Cloud Infrastructure (ABCI) which is set to become the fastest supercomputer system in Japan Read more…

By John Russell

HPC Chips – A Veritable Smorgasbord?

October 10, 2017

For the first time since AMD's ill-fated launch of Bulldozer the answer to the question, 'Which CPU will be in my next HPC system?' doesn't have to be 'Whichever variety of Intel Xeon E5 they are selling when we procure'. Read more…

By Dairsie Latimer

Delays, Smoke, Records & Markets – A Candid Conversation with Cray CEO Peter Ungaro

October 5, 2017

Earlier this month, Tom Tabor, publisher of HPCwire and I had a very personal conversation with Cray CEO Peter Ungaro. Cray has been on something of a Cinderell Read more…

By Tiffany Trader & Tom Tabor

Intel Debuts Programmable Acceleration Card

October 5, 2017

With a view toward supporting complex, data-intensive applications, such as AI inference, video streaming analytics, database acceleration and genomics, Intel i Read more…

By Doug Black

OLCF’s 200 Petaflops Summit Machine Still Slated for 2018 Start-up

October 3, 2017

The Department of Energy’s planned 200 petaflops Summit computer, which is currently being installed at Oak Ridge Leadership Computing Facility, is on track t Read more…

By John Russell

US Exascale Program – Some Additional Clarity

September 28, 2017

The last time we left the Department of Energy’s exascale computing program in July, things were looking very positive. Both the U.S. House and Senate had pas Read more…

By Alex R. Larzelere

How ‘Knights Mill’ Gets Its Deep Learning Flops

June 22, 2017

Intel, the subject of much speculation regarding the delayed, rewritten or potentially canceled “Aurora” contract (the Argonne Lab part of the CORAL “ Read more…

By Tiffany Trader

Reinders: “AVX-512 May Be a Hidden Gem” in Intel Xeon Scalable Processors

June 29, 2017

Imagine if we could use vector processing on something other than just floating point problems.  Today, GPUs and CPUs work tirelessly to accelerate algorithms Read more…

By James Reinders

NERSC Scales Scientific Deep Learning to 15 Petaflops

August 28, 2017

A collaborative effort between Intel, NERSC and Stanford has delivered the first 15-petaflops deep learning software running on HPC platforms and is, according Read more…

By Rob Farber

Oracle Layoffs Reportedly Hit SPARC and Solaris Hard

September 7, 2017

Oracle’s latest layoffs have many wondering if this is the end of the line for the SPARC processor and Solaris OS development. As reported by multiple sources Read more…

By John Russell

US Coalesces Plans for First Exascale Supercomputer: Aurora in 2021

September 27, 2017

At the Advanced Scientific Computing Advisory Committee (ASCAC) meeting, in Arlington, Va., yesterday (Sept. 26), it was revealed that the "Aurora" supercompute Read more…

By Tiffany Trader

Google Releases Deeplearn.js to Further Democratize Machine Learning

August 17, 2017

Spreading the use of machine learning tools is one of the goals of Google’s PAIR (People + AI Research) initiative, which was introduced in early July. Last w Read more…

By John Russell

GlobalFoundries Puts Wind in AMD’s Sails with 12nm FinFET

September 24, 2017

From its annual tech conference last week (Sept. 20), where GlobalFoundries welcomed more than 600 semiconductor professionals (reaching the Santa Clara venue Read more…

By Tiffany Trader

Graphcore Readies Launch of 16nm Colossus-IPU Chip

July 20, 2017

A second $30 million funding round for U.K. AI chip developer Graphcore sets up the company to go to market with its “intelligent processing unit” (IPU) in Read more…

By Tiffany Trader

Leading Solution Providers

Amazon Debuts New AMD-based GPU Instances for Graphics Acceleration

September 12, 2017

Last week Amazon Web Services (AWS) streaming service, AppStream 2.0, introduced a new GPU instance called Graphics Design intended to accelerate graphics. The Read more…

By John Russell

Nvidia Responds to Google TPU Benchmarking

April 10, 2017

Nvidia highlights strengths of its newest GPU silicon in response to Google's report on the performance and energy advantages of its custom tensor processor. Read more…

By Tiffany Trader

EU Funds 20 Million Euro ARM+FPGA Exascale Project

September 7, 2017

At the Barcelona Supercomputer Centre on Wednesday (Sept. 6), 16 partners gathered to launch the EuroEXA project, which invests €20 million over three-and-a-half years into exascale-focused research and development. Led by the Horizon 2020 program, EuroEXA picks up the banner of a triad of partner projects — ExaNeSt, EcoScale and ExaNoDe — building on their work... Read more…

By Tiffany Trader

Delays, Smoke, Records & Markets – A Candid Conversation with Cray CEO Peter Ungaro

October 5, 2017

Earlier this month, Tom Tabor, publisher of HPCwire and I had a very personal conversation with Cray CEO Peter Ungaro. Cray has been on something of a Cinderell Read more…

By Tiffany Trader & Tom Tabor

Cray Moves to Acquire the Seagate ClusterStor Line

July 28, 2017

This week Cray announced that it is picking up Seagate's ClusterStor HPC storage array business for an undisclosed sum. "In short we're effectively transitioning the bulk of the ClusterStor product line to Cray," said CEO Peter Ungaro. Read more…

By Tiffany Trader

Intel Launches Software Tools to Ease FPGA Programming

September 5, 2017

Field Programmable Gate Arrays (FPGAs) have a reputation for being difficult to program, requiring expertise in specialty languages, like Verilog or VHDL. Easin Read more…

By Tiffany Trader

IBM Advances Web-based Quantum Programming

September 5, 2017

IBM Research is pairing its Jupyter-based Data Science Experience notebook environment with its cloud-based quantum computer, IBM Q, in hopes of encouraging a new class of entrepreneurial user to solve intractable problems that even exceed the capabilities of the best AI systems. Read more…

By Alex Woodie

Intel, NERSC and University Partners Launch New Big Data Center

August 17, 2017

A collaboration between the Department of Energy’s National Energy Research Scientific Computing Center (NERSC), Intel and five Intel Parallel Computing Cente Read more…

By Linda Barney

  • arrow
  • Click Here for More Headlines
  • arrow
Share This