Power8 with NVLink Coming to the Nimbix Cloud

By Tiffany Trader

October 6, 2016

Starting later this month, HPC professionals and data scientists wishing to try out NVLink’d Nvidia Pascal P100 GPUs won’t have to spend upwards of $100,000 on NVIDIA’s DGX-1 server or fork over about half that for IBM’s Power8 server with NVLink and four Pascal GPUs. Soon they’ll be able to get the power of Pascal in a public cloud.

On Wednesday (Oct. 5), the Dallas, Texas-based cloud provider Nimbix revealed that it was adding IBM Power S822LC for HPC systems (codenamed “Minsky”) to its heterogeneous HPC cloud platform. Target markets include high-performance computing, data analytics, in-memory databases, and machine learning.

“We are definitely the first public cloud to deploy the Minsky technology and one of the first to deploy Power8 in a high performance high-scalability setting,” said Leo Reiter, CTO and vice president of software engineering at Nimbix. “Obviously what was really interesting on Minksky in addition to Power8 was the Pascal GPUs and we’ve integrated that with our JARVICE platform so it’s a seamless experience for both end users and developers.”

Unveiled by IBM last month, the new Power8 with NVLink processor features 10 cores running up to 3.26 GHz. The processors have higher memory bandwidth than x86 CPUs at 115 GB/s and can have as much as half a terabyte of system memory per socket. There are larger caches per core inside the Power8 processor, and this coupled with the faster cores and memory bandwidth leads to higher application performance and throughput.

The NVIDIA Tesla P100 for NVLink-optimized servers is Nvidia’s most performant GPU yet, delivering a whopping 5.3 teraflops of double-precision performance, 10.6 teraflops of single-precision, and 21 teraflops of half-precision. The accelerator card includes 16 gigabytes of the HBM2 stacked memory with an on-GPU memory bandwidth of 720 GB/s. The Tesla P100 with NVLink GPU in the SXM2 (Mezzanine) form factor, currently only shipping in the DGX-1 and the Minsky platform, delivers 13 percent more raw compute performance than the PCIe variant due to the higher TDP (300 watts versus 250 watts).

IBM.POWER8.NVLINKCrucially for many users that Nimbix is targeting with the new hardware, the Minsky platform provides high-bandwidth NVLink connections between the CPU and the GPUs and from GPU to GPU. IBM says the NVLink optimized Power8 servers “enable data to flow 5x faster” than on a comparable x86-based system.

Along with its support for HPC and deep learning workflows, Nimbix says adoption of GPU-accelerated databases is advancing quickly. “Accelerated analytics, like in-memory databases, benefit so much from having the Pascal GPUs as well as the high performance link between them to the point where some customers are getting multiple times performance boost for advanced queries,” says Reiter.

Nimbix is working with Kinetica and MapD to facilitate the use of the NVLink optimized Power servers for database acceleration. Reiter says Kinetica has the ability to scale horizontally across multiple systems so it’s not just being able to take advantage of the GPUs on a single box but to scale across using a high performance fabric like the one at Nimbix.

“So you have both the high density in each chassis where you are going to have four Pascal GPUs with this high performance link, a lot of host memory but then also being able to scale that horizontally to dozens of machines at the same time to be able to do these accelerated databases,” says Reiter.

From its start in 2010, Nimbix has focused on high-performance heterogeneous cloud computing. “While it’s true that the market is heavily tilted toward Intel in terms of the system architecture, we already have the capabilities of running heterogeneous compute, both accelerators as well as central processors, so it wasn’t a technology challenge for us to deploy a non-Intel architecture into our existing cloud,” explains Reiter.

“We’re not requiring that people embrace the system architecture change. We’re not selling architecture here, what we’re selling is turnkey workflows that happen to run in a more optimized way when they’re hooked into the right resources.”

Reiter notes that they’ve been working with Nvidia for a while now and they have multiple datacenters outfitted with Tesla K80s and other Nvidia chips.

Now they’re also partnering with IBM, whom they say has provided and continues to provide a lot of support. “They are extremely motivated to speed Power,” says Reiter, “not just in HPC, but in cloud specifically. And there are a couple reasons. One is that the traction in public cloud for Power is almost non-existent, relatively speaking, but more importantly, even a lot of their customers who are buying Power clusters are asking them for an actual cloud bursting strategy. It’s not just for capacity, but it’s for a lot of these bids now, people are looking at emerging technology in the bids.”

The secret sauce of the Nimbix cloud is the JARVICE container runtime. “The native execution model for JARVICE is containers running on bare metal, so there are no virtual machines,” says Reiter. “There are containers but these are custom-built containers, not Docker. There was too much overhead and complexity and performance loss with getting Docker to run, especially for tightly-coupled, high performance workloads, so we designed our container technology from the ground up, and [with our PushtoCompute technology] we accept ordinary Docker containers as input and convert them on the fly to run natively on the JARVICE platform.”

Asked if they would ever productize JARVICE outside of the Nimbix cloud, Reiter said they are happy to discuss that with anyone who is interested but do not have an immediate play to offer it as an off-the-shelf software product. That said, Nimbix does have select datacenter customers who are trying out the software.

Nimbix’s specialty is enabling turnkey cloud via a software-as-a-service delivery model. It’s not for the user looking to spin up virtual clouds.

“IaaS public cloud is fine for dev-tests of single machine instances,” says Reiter. “Sometimes you want to test out some code and see if it works and that’s great. But when we’re talking about deploying tightly-coupled workflows at scale, deploying the software and tuning the software is extremely complicated.”

“What customers enjoy on Nimbix is they look for the workflow they want to run, they click on it, they specify whatever parameters are relevant to that workflow and they click submit and then their data comes back processed the way they want it to without having to care about ‘How am I going to scale this? How am I going to install it? Am I running the right version? Do I have the right libraries installed?’ JARVICE takes care of all of that – and it’s extensible through technologies like PushtoCompute to enable the onboarding of more and more functionality.”

Every machine in the Nimbix cloud is InfiniBand connected via a Mellanox EDR InfiniBand spine and FDR InfiniBand to the compute nodes. JARVICE also employs distributed block storage and distributed storage over InfiniBand, plus 20GB Ethernet (bonded 10GB) for accessing the internet.

Nimbix expects to have customers using the Minsky platform publicly by the end of this month and prior to that will be conducting benchmarking tests with early-access customers. Initially, demand will likely outstrip availability, and Nimbix says it’s already planning the next step of the expansion, essentially taking orders as soon as IBM can ship.

“The line is there and it’s only going to get bigger,” observes Reiter. “We’re excited to be able service these customers and position ourselves to service more in the future.”

Pricing has not yet been disclosed, but Minsky will be offered in single-GPU and quad-GPU units – via subscription or pay-as-you-go.

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industy updates delivered to you every week!

What’s New in HPC Research: Natural Gas, Precision Agriculture, Neural Networks and More

December 6, 2019

In this bimonthly feature, HPCwire highlights newly published research in the high-performance computing community and related domains. From parallel programming to exascale to quantum computing, the details are here. Read more…

By Oliver Peckham

On the Spack Track @SC19

December 5, 2019

At the annual supercomputing conference, SC19 in Denver, Colorado, there were Spack events each day of the conference. As a reflection of its grassroots heritage, nine sessions were planned by more than a dozen thought leaders from seven organizations, including three U.S. national Department of Energy (DOE) laboratories and Sylabs... Read more…

By Elizabeth Leake

Intel’s New Hyderabad Design Center Targets Exascale Era Technologies

December 3, 2019

Intel's Raja Koduri was in India this week to help launch a new 300,000 square foot design and engineering center in Hyderabad, which will focus on advanced computing technologies for the AI and exascale era. "Over th Read more…

By Tiffany Trader

AWS Debuts 7nm 2nd-Gen Graviton Arm Processor

December 3, 2019

The “x86 Big Bang,” in which market dominance of the venerable Intel CPU has exploded into fragments of processor options suited to varying workloads, has now encompassed CPUs offered by the leading public cloud serv Read more…

By Doug Black

Medical Imaging Gets an AI Boost

December 3, 2019

AI technologies incorporated into diagnostic imaging tools have proven useful in eliminating confirmation bias, often outperforming human clinicians who may bring their own prejudices. Another issue slowing progress is t Read more…

By George Leopold

AWS Solution Channel

Making High Performance Computing Affordable and Accessible for Small and Medium Businesses with HPC on AWS

High performance computing (HPC) brings a powerful set of tools to a broad range of industries, helping to drive innovation and boost revenue in finance, genomics, oil and gas extraction, and other fields. Read more…

IBM Accelerated Insights

AI Needs Intelligent HPC infrastructure

Artificial Intelligence (AI) has revolutionized entire industries and enables humanity to solve some of the most daunting challenges. To accomplish this, it requires massive amounts of data from heterogeneous sources that is processed it new ways that differs significantly from HPC applications. Read more…

Ride on the Wild Side – Squyres SC19 Mars Rovers Keynote

December 2, 2019

Reminding us of the deep and enabling connection between HPC and modern science is an important part of the SC Conference mission. And yes, HPC is a science itself. At SC19, Steve Squyres’ opening keynote recounting th Read more…

By John Russell

On the Spack Track @SC19

December 5, 2019

At the annual supercomputing conference, SC19 in Denver, Colorado, there were Spack events each day of the conference. As a reflection of its grassroots heritage, nine sessions were planned by more than a dozen thought leaders from seven organizations, including three U.S. national Department of Energy (DOE) laboratories and Sylabs... Read more…

By Elizabeth Leake

Intel’s New Hyderabad Design Center Targets Exascale Era Technologies

December 3, 2019

Intel's Raja Koduri was in India this week to help launch a new 300,000 square foot design and engineering center in Hyderabad, which will focus on advanced com Read more…

By Tiffany Trader

AWS Debuts 7nm 2nd-Gen Graviton Arm Processor

December 3, 2019

The “x86 Big Bang,” in which market dominance of the venerable Intel CPU has exploded into fragments of processor options suited to varying workloads, has n Read more…

By Doug Black

Ride on the Wild Side – Squyres SC19 Mars Rovers Keynote

December 2, 2019

Reminding us of the deep and enabling connection between HPC and modern science is an important part of the SC Conference mission. And yes, HPC is a science its Read more…

By John Russell

NSCI Update – Adapting to a Changing Landscape

December 2, 2019

It was November of 2017 when we last visited the topic of the National Strategic Computing Initiative (NSCI). As you will recall, the NSCI was started with an Executive Order (E.O. No. 13702), that was issued by President Obama in July of 2015 and was followed by a Strategic Plan that was released in July of 2016. The question for November of 2017... Read more…

By Alex R. Larzelere

Tsinghua University Racks Up Its Ninth Student Cluster Championship Win at SC19

November 27, 2019

Tsinghua University has done it again. At SC19 last week, the eight-time gold medal-winner team took home the top prize in the 2019 Student Cluster Competition Read more…

By Oliver Peckham

SC19: IBM Changes Its HPC-AI Game Plan

November 25, 2019

It’s probably fair to say IBM is known for big bets. Summit supercomputer – a big win. Red Hat acquisition – looking like a big win. OpenPOWER and Power processors – jury’s out? At SC19, long-time IBMer Dave Turek sketched out a different kind of bet for Big Blue – a small ball strategy, if you’ll forgive the baseball analogy... Read more…

By John Russell

How the Gordon Bell Prize Winners Used Summit to Illuminate Transistors

November 22, 2019

At SC19, the Association for Computing Machinery (ACM) awarded the prestigious Gordon Bell Prize to the Swiss Federal Institute of Technology (ETH) Zurich. The Read more…

By Oliver Peckham

Using AI to Solve One of the Most Prevailing Problems in CFD

October 17, 2019

How can artificial intelligence (AI) and high-performance computing (HPC) solve mesh generation, one of the most commonly referenced problems in computational engineering? A new study has set out to answer this question and create an industry-first AI-mesh application... Read more…

By James Sharpe

Cray Wins NNSA-Livermore ‘El Capitan’ Exascale Contract

August 13, 2019

Cray has won the bid to build the first exascale supercomputer for the National Nuclear Security Administration (NNSA) and Lawrence Livermore National Laborator Read more…

By Tiffany Trader

DARPA Looks to Propel Parallelism

September 4, 2019

As Moore’s law runs out of steam, new programming approaches are being pursued with the goal of greater hardware performance with less coding. The Defense Advanced Projects Research Agency is launching a new programming effort aimed at leveraging the benefits of massive distributed parallelism with less sweat. Read more…

By George Leopold

D-Wave’s Path to 5000 Qubits; Google’s Quantum Supremacy Claim

September 24, 2019

On the heels of IBM’s quantum news last week come two more quantum items. D-Wave Systems today announced the name of its forthcoming 5000-qubit system, Advantage (yes the name choice isn’t serendipity), at its user conference being held this week in Newport, RI. Read more…

By John Russell

Ayar Labs to Demo Photonics Chiplet in FPGA Package at Hot Chips

August 19, 2019

Silicon startup Ayar Labs continues to gain momentum with its DARPA-backed optical chiplet technology that puts advanced electronics and optics on the same chip Read more…

By Tiffany Trader

SC19: IBM Changes Its HPC-AI Game Plan

November 25, 2019

It’s probably fair to say IBM is known for big bets. Summit supercomputer – a big win. Red Hat acquisition – looking like a big win. OpenPOWER and Power processors – jury’s out? At SC19, long-time IBMer Dave Turek sketched out a different kind of bet for Big Blue – a small ball strategy, if you’ll forgive the baseball analogy... Read more…

By John Russell

Cray, Fujitsu Both Bringing Fujitsu A64FX-based Supercomputers to Market in 2020

November 12, 2019

The number of top-tier HPC systems makers has shrunk due to a steady march of M&A activity, but there is increased diversity and choice of processing compon Read more…

By Tiffany Trader

Crystal Ball Gazing: IBM’s Vision for the Future of Computing

October 14, 2019

Dario Gil, IBM’s relatively new director of research, painted a intriguing portrait of the future of computing along with a rough idea of how IBM thinks we’ Read more…

By John Russell

Leading Solution Providers

ISC 2019 Virtual Booth Video Tour

CRAY
CRAY
DDN
DDN
DELL EMC
DELL EMC
GOOGLE
GOOGLE
ONE STOP SYSTEMS
ONE STOP SYSTEMS
PANASAS
PANASAS
VERNE GLOBAL
VERNE GLOBAL

Intel Debuts New GPU – Ponte Vecchio – and Outlines Aspirations for oneAPI

November 17, 2019

Intel today revealed a few more details about its forthcoming Xe line of GPUs – the top SKU is named Ponte Vecchio and will be used in Aurora, the first plann Read more…

By John Russell

Kubernetes, Containers and HPC

September 19, 2019

Software containers and Kubernetes are important tools for building, deploying, running and managing modern enterprise applications at scale and delivering enterprise software faster and more reliably to the end user — while using resources more efficiently and reducing costs. Read more…

By Daniel Gruber, Burak Yenier and Wolfgang Gentzsch, UberCloud

Dell Ramps Up HPC Testing of AMD Rome Processors

October 21, 2019

Dell Technologies is wading deeper into the AMD-based systems market with a growing evaluation program for the latest Epyc (Rome) microprocessors from AMD. In a Read more…

By John Russell

AMD Launches Epyc Rome, First 7nm CPU

August 8, 2019

From a gala event at the Palace of Fine Arts in San Francisco yesterday (Aug. 7), AMD launched its second-generation Epyc Rome x86 chips, based on its 7nm proce Read more…

By Tiffany Trader

SC19: Welcome to Denver

November 17, 2019

A significant swath of the HPC community has come to Denver for SC19, which began today (Sunday) with a rich technical program. As is customary, the ribbon cutt Read more…

By Tiffany Trader

When Dense Matrix Representations Beat Sparse

September 9, 2019

In our world filled with unintended consequences, it turns out that saving memory space to help deal with GPU limitations, knowing it introduces performance pen Read more…

By James Reinders

With the Help of HPC, Astronomers Prepare to Deflect a Real Asteroid

September 26, 2019

For years, NASA has been running simulations of asteroid impacts to understand the risks (and likelihoods) of asteroids colliding with Earth. Now, NASA and the European Space Agency (ESA) are preparing for the next, crucial step in planetary defense against asteroid impacts: physically deflecting a real asteroid. Read more…

By Oliver Peckham

Cerebras to Supply DOE with Wafer-Scale AI Supercomputing Technology

September 17, 2019

Cerebras Systems, which debuted its wafer-scale AI silicon at Hot Chips last month, has entered into a multi-year partnership with Argonne National Laboratory and Lawrence Livermore National Laboratory as part of a larger collaboration with the U.S. Department of Energy... Read more…

By Tiffany Trader

  • arrow
  • Click Here for More Headlines
  • arrow
Do NOT follow this link or you will be banned from the site!
Share This