Visit additional Tabor Communication Publications
June 05, 2008
The cloud computing meme is permeating practically all areas of computing these days, including HPC. Will the cloud replace the grid as the new paradigm for delivering high performance computing? To be fair, it's not that grid computing never delivered; it just never reached escape velocity for the HPC market.
With the emergence of commercial solutions, like Amazon EC2, Google App Engine, Sun's Network.com, and IBM's Deep Computing on Demand, cloud-based utility computing is starting to look a lot more mainstream. Not wanting to be left out of "The Next Big Thing," HP, Dell, Microsoft, Yahoo, and others are all jumping onto the cloud-wagon -- if you'll pardon the mixed metaphor.
So will the cloud swallow HPC just as commodity computing machines did ten years ago? Probably, but it's too soon too tell how fast that will occur. The technical challenges -- networking bandwidth and latency, compute performance, and software standards -- are still an issue for HPC but are quickly being solved for enterprise computing. And the business model for large scale utility computing is still being worked out.
Cloud computing maven Nick Carr has some ideas about that. In a recent post on his blog, Carr says organizations like Amazon or Google have the edge right now. His hypothesis is that companies that have already built an ultra-scale compute infrastructure for their low-margin retail or advertising business can leverage that investment into a higher-margin utility computing business. "For Amazon, running a cloud computing service is core to its business in a way that it isn't for, say, IBM, Sun, or HP," explains Carr.
How does that translate to HPC? Ultra-scale platforms like Google and Amazon are not really set up for HPC apps at this point. The only large scale supercomputing infrastructures are owned by governments, TeraGrid in the U.S and DEISA in Europe being the best examples. They have no way to deliver all that compute and storage capacity as a commercial utility solution and, being research-oriented, have no mandate to do so.
Outside of the cloud model, individual HPC systems can be farmed out. The DOE, for example, shares its high-end supercomputers with industry (and academia) via its highly regarded INCITE program, but not as a commercial service. The lucky few companies that win the INCITE lottery get to use the cutting-edge supers for high-end industrial research, but not for day-to-day computing. While government-industry HPC partnerships have become rather common, the model isn't geared for production work.
In the commercial arena, IBM's Deep Computing Capacity On-Demand (DCCoD) rents out Blue Gene cycles as well as capacity on less exotic platforms from its DCCoD centers. And Sun's Network.com lets customers buy compute time by the CPU-hour on a modest-sized x86 cluster, while also offering access to a handful of HPC application suites. The long-term viability of this model is still a question.
More in the Nick Carr model of doubling up on in-house computing resources, the Computational Research Laboratories (CRL) in Pune, India, is offering up Eka, its new 117.8 teraflop supercomputer for commercial use. As the number four system on the current TOP500 list, Eka is the only privately owned supercomputer in the top 10. CRL itself is owned and operated by the Tata Group, and according to a Financial Express report, the machine will be used by the group's Tata Motors (automotive) and Tata Elxsi (product design) subsidiaries. By also offering the $30 million machine as a supercomputing service platform, Tata intends to get the most return on its investment.
An IEEE Spectrum article published today talked about Tata's CRL HPC business strategy:
True to India's software and services tech culture, rather than try to outdo Cray, IBM, Hewlett-Packard, or Silicon Graphics at designing and selling supercomputers, CRL will provide end-to-end supercomputing services—renting computer time, adapting and fine-tuning applications, and offering analytical services. Today Eka is testing more than 15 applications for customers, and the company is in talks with several clients from the automobile, aerospace, financial, oil and gas exploration, and life sciences sectors, including aerospace giants Boeing, Embraer-Empresa Brasileira de Aeronautica, and Airbus.
Even if this arrangement proves to be workable for Tata, the model would be hard to reproduce. How many commercial organizations can afford to buy supercomputing resources at a scale that would serve a reasonable number of HPC renters and offer the kind of software support that CRL is intending to provide? Besides Tata, I can't think of anyone else.
Which brings us back to cloud computing. The way I envision HPC moving over to the cloud in mass is when the aggregate performance of the systems is so great that the lack of efficiency won't really matter. These systems will be able to "waste" compute, storage and network resources because economies of scale will render monolithic machines way too expensive. We'll know that has happened when NCAR is running their climate models on Google EarthSim and Boeing is designing airplanes on Amazon WindTunnel.
Posted by Michael Feldman - June 04, 2008 @ 9:00 PM, Pacific Daylight Time
Michael Feldman is the editor of HPCwire.
No Recent Blog Comments
The Xeon Phi coprocessor might be the new kid on the high performance block, but out of all first-rate kickers of the Intel tires, the Texas Advanced Computing Center (TACC) got the first real jab with its new top ten Stampede system.We talk with the center's Karl Schultz about the challenges of programming for Phi--but more specifically, the optimization...
Although Horst Simon was named Deputy Director of Lawrence Berkeley National Laboratory, he maintains his strong ties to the scientific computing community as an editor of the TOP500 list and as an invited speaker at conferences.
Supercomputing veteran, Bo Ewald, has been neck-deep in bleeding edge system development since his twelve-year stint at Cray Research back in the mid-1980s, which was followed by his tenure at large organizations like SGI and startups, including Scale Eight Corporation and Linux Networx. He has put his weight behind quantum company....
May 16, 2013 |
When it comes to cloud, long distances mean unacceptably high latencies. Researchers from the University of Bonn in Germany examined those latency issues of doing CFD modeling in the cloud by utilizing a common CFD and its utilization in HPC instance types including both CPU and GPU cores of Amazon EC2.
May 15, 2013 |
Supercomputers at the Department of Energy’s National Energy Research Scientific Computing Center (NERSC) have worked on important computational problems such as collapse of the atomic state, the optimization of chemical catalysts, and now modeling popping bubbles.
May 10, 2013 |
Program provides cash awards up to $10,000 for the best open-source end-user applications deployed on 100G network.
May 09, 2013 |
The Japanese government has revealed its plans to best its previous K Computer efforts with what they hope will be the first exascale system...
May 08, 2013 |
For engineers looking to leverage high-performance computing, the accessibility of a cloud-based approach is a powerful draw, but there are costs that may not be readily apparent.
05/10/2013 | Cleversafe, Cray, DDN, NetApp, & Panasas | From Wall Street to Hollywood, drug discovery to homeland security, companies and organizations of all sizes and stripes are coming face to face with the challenges – and opportunities – afforded by Big Data. Before anyone can utilize these extraordinary data repositories, however, they must first harness and manage their data stores, and do so utilizing technologies that underscore affordability, security, and scalability.
04/15/2013 | Bull | “50% of HPC users say their largest jobs scale to 120 cores or less.” How about yours? Are your codes ready to take advantage of today’s and tomorrow’s ultra-parallel HPC systems? Download this White Paper by Analysts Intersect360 Research to see what Bull and Intel’s Center for Excellence in Parallel Programming can do for your codes.
In this demonstration of SGI DMF ZeroWatt disk solution, Dr. Eng Lim Goh, SGI CTO, discusses a function of SGI DMF software to reduce costs and power consumption in an exascale (Big Data) storage datacenter.
The Cray CS300-AC cluster supercomputer offers energy efficient, air-cooled design based on modular, industry-standard platforms featuring the latest processor and network technologies and a wide range of datacenter cooling requirements.