South Africa Maps HPC Future as SKA Plans Take Shape

By Elizabeth Leake, STEM-Trek -- Photography by Lawrette McFarlane

March 1, 2016

The South African Center for High Performance Computing’s (SA-CHPC) Ninth Annual National Meeting was held Nov. 30 – Dec. 4, 2015, at the Council for Scientific and Industrial Research (CSIR) International Convention Center in Pretoria, SA. The award-winning venue was the perfect location to host what has become a popular industry, regional and educational showcase.

Exascale in the 2020s

With the Square Kilometer Array (SKA) being built in the great Karoo region, implications for SA and the HPC industry have captured the attention of a broad range of stakeholders. SKA will be the world’s biggest radio telescope, and the most ambitious technology project ever funded. With an expected 50-year lifespan, SKA Phase One construction is scheduled to begin in 2018, and early science and data generation will follow by 2020.

With only a few years in which to prepare for SKA, it’s not surprising that the CHPC conference has begun to feature in-depth data science and network infrastructure content, in addition to the meeting’s traditional HPC tutorials, workshops, plenaries, and student programs. Once again, the Southern African Development Community (SADC) HPC and Industry Forums were co-located with the CHPC meeting. An increased number of conference attendees from data science and high-speed network occupations added diversity in terms of gender, discipline and nationalities represented. All things considered, the annual CHPC meeting provides a wealth of learning opportunities, and brings professional networking to an emerging region of the HPC world.

Just how resource-intensive will SKA be?

Peter Braam (SKA Cambridge University, Parallel Scientific)
Peter Braam (SKA Cambridge University, Parallel Scientific)

It’s safe to say that SKA will set the “Big Data” curve. In his plenary address, Peter Braam (SKA Cambridge University, Parallel Scientific) described the range of instrumentation that will support the project, and how a combination of dedicated and cloud-enabled resources will fulfill its varied and complex missions. As for data, by 2020 SKA’s central signal processing array—50 times more sensitive than any current radio instrument—will generate 1 exabytes per day, and 100 exabytes per day by 2028. Its imaging function will produce 400 terabytes daily for worldwide consumption by 2020, and 10,000 petabytes per day by 2028. As for computational processing power, it will require 300 petaflops (archiving 1 exabyte) by 2020 and 10 times more by 2028.

Rudolph Pienaar (Boston Children’s Hospital, Harvard) noted that even before SKA data enters the infostream, by 2019 general network traffic will explode as a larger percentage of the population has access to networks; with an additional 30 percent more networked devices per capita and many more applications that are increasingly resource-intensive. Global IP traffic will reach 1.1 zettabytes per year in 2016, or 88.4 exabytes (one billion gigabytes) per month. By 2019, global IP traffic will reach 2.0 zettabytes per year, or 168 exabytes per month.

Most of the Karoo region and surrounding areas currently lack high-speed networks, and a reliable electrical supply. However, with ten wealthier countries and more than 100 organizations behind the project, SA expects to have the support it will need to host SKA. Additionally, many of SADC’s 15 member states will benefit from an improved power grid and access to high-speed networks, in addition to slipstream advantages such as workforce development opportunities and access to tools that support international research engagement.

CHPC National Meeting

Happy Sithole (CHPC/CSIR-SA)
Happy Sithole (CHPC/CSIR-SA)

The 2015 meeting was called to order by General Chair Janse Van Rensburg (CHPC) who introduced Director General Phil Mjwara (SA Dept. of Science & Technology). Dr. Mjwara expressed the vital role HPC will play in South Africa’s future. “A greater investment in cyberinfrastructure (CI) and supportive human capital development are essential for good governance, sustainability and commerce,” he said.

Mjwara shared 2015 highlights, including a milestone partnership between CHPC and SANRen. The agreement, reached in April, 2015, allows CHPC to engage with 155 tier-2 sites around the world that access the Large Hadron Collider at CERN in Geneva, Switzerland.

SA-CHPC New System FeaturesCHPC Director Happy Sithole described how their Cape Town center has grown since launching in 2007 when it supported 15 researchers with 2.5 teraflops of computational power. “That was a good start for South Africa, and CHPC was the only center on the continent at the time,” said Sithole. Since then, CHPC has expanded to meet demands. At the time of the 2015 conference, CHPC supported 700 researchers with 7,000 cores and 64 teraflops. “But, we could be fully subscribed with our three largest research projects,” he added.

Sithole was pleased to announce the addition of a new Dell system that will be operational in early 2016. Launched in two phases, the first is expected to achieve 700 teraflops and the second will add an additional 300 teraflops with GPU acceleration. “This brings a total of one petaflops of computational power to the region,” he added. Thirty percent of the system’s cycles will be available for lease by CSIR industry partners.

Thomas Sterling (Indiana University)
Thomas Sterling (Indiana University)

The student poster and cluster competition awards were presented Thursday evening following a plenary address by Thomas Sterling (Indiana University). Sterling is executive associate director and chief scientist at the Center for Research in Extreme Scale Technologies (CREST) and is best known for his pioneering work in commodity clusters as “the father of Beowulf.” He noted the accelerated progress being made by CHPC and pioneer projects in region, and credited Sithole for blazing the trail on behalf of industry, science and education.

“The new CHPC Dell and Mellanox system will provide world-class compute capability to prepare the emerging southern African region for exascale computing in the next decade,” said Sterling.

Plenary Highlights

The opening plenary address by Merle Giles (National Center for Supercomputing Applications, U.S.) was titled “HPC-Enabled Innovation and Transformational Science & Engineering: The Role of CI.” As the leader of NCSA’s private sector program, Giles explained HPC’s macro and micro economic impact on the economy, with the federal investment defining macro-, and micro-economic impact coming from universities and industry.

Giles described the “R&D Valley of Death,” or the funding gap that exists between the theoretical and basic research led by universities (or startups), and the production-commercialization phase effectuated by industry. “Somewhere in the middle, before optimization and ‘robustification,’ great ideas tend to crash and burn for lack of funding,” he said. “When this happens, it adversely affects how the public perceives science and technology spending, in general,” he added. Greater emphases on investment and tech transfer will help bridge this gap.

Merle Giles (NCSA Private Sector Program - U.S.)
Merle Giles (NCSA Private Sector Program – U.S.)

Giles said President Obama initiated an important discussion in January, 2015, during his state of the union address. By explaining how precision medicine and the curing of common diseases are enabled by fast computers and data science, he demystified the concept so that average citizens could understand why HPC is important to them. He then announced the nation’s strategic computing initiative in August. “Through this steady dialogue, Obama is laying the groundwork for future support of a greater public investment in advanced CI,” Giles said.

He added that all sectors—public, private and industry—must share the responsibility for preparing the workforce. “And, by the way, coding is no longer the new literacy; modeling and the ability to master data are more important as HPC becomes more accessible, and the long tail of science engages a broader range of practitioners,” he added. He concluded with facts from an International Data Corporation study that found for every dollar invested in HPC, $515 dollars in revenue and $43 dollars of profit are generated. Feasibility should be addressed, but it’s often overlooked. No single company, agency or government can fund everything. A sustainable program requires collaboration and a holistic approach to spending.

Peter Braam (SKA Cambridge University, Parallel Scientific) presented “HPC Stack—Scientific Computing Meets Cloud.” Braam explained how cloud-augmented cluster services satisfy a greater spectrum of administrative and production applications than earth-bound HPC can on its own. By provisioning clusters with containers, or cloud-enabled virtual machines (VMs), it’s possible to facilitate core services, such as identity management, storage, scheduling, and monitoring. “We discovered there are generally two groups of users: people performing operations, and developers who are experimenting in pursuit of discovery. To support the latter, we realized it’s more feasible to deploy 100 small virtual cloud-enabled environments than 100 HPC systems.” Cambridge, CHPC and Canonical Ltd are pursuing this study.

Simon Hodson (CODATA - France)
Simon Hodson (CODATA – France)

Thursday’s plenary was delivered by CODATA Executive Director Simon Hodson, and was titled “Mobilizing the data revolution: CODATA’s work on data policy, data science and capacity building.”

Hodson suggested that a 2012 report by The Royal Society titled “Science as an open enterprise” made a profound impact on the industry when it stated that data underpinning research findings must be openly available. “Transparency, credibility and reproducibility are the keys to gaining public confidence; data must therefore be accessible, assessable, intelligible, usable, and reusable,” he said.

The European Union’s Horizon 2020 project included, word-for-word, excerpts from the report that suggested principal investigators are guilty of scientific misconduct if their data aren’t accessible. Most European and U.S. agencies now require data generated by research they fund to be liberated, and many have issued policies regarding open research data. Leading industry journals and digital repositories have also adopted the practice, including Dryad, GenBank and TreeBASE.

Hodson addressed data citation best practices, and financing strategies to support sustainable data repositories. He presented the findings of a joint declaration of data citation principles by the CODATA-ICSTI Task Group on Data Citation Standards and Practices and FORCE11 (Future of eScience Communications and Scholarship). “Make data available in a usable way, so they can be recognized and properly credited,” he said, referring to a CODATA paper published in the Journal of Data Science titled “Out of Cite, Out of Mind: The Current State of Practice, Policy, and Technology for the Citation of Data.”

Hodson announced a series of 2016-17 Data Science Summer Schools that are co-sponsored by CODATA and the Research Data Alliance (RDA). The two-week programs will cover software and data carpentry, data stewardship, curation, visualization, and more. CODATA-RDA earmarked $60,000 for travel funding for registrants who lack institutional support. The first will be held in Trieste, Italy on August 1-12, 2016. The International Center for Theoretical Physics (ICTP) will support up to 120 students, and 30,000 Euros are allocated for student travel. Students from low and middle-income countries (LMICs) will be given priority, and additional sponsors are welcome.

Rudolph Pienaar (Boston Children's Hospital, Harvard)
Rudolph Pienaar (Boston Children’s Hospital, Harvard)

Health data concerns were addressed by Rudolph Pienaar (Boston Children’s Hospital) with his presentation titled “Web 2.0 and beyond: Leveraging Web Technologies as Middleware in Healthcare and High Performance Compute Clusters; Data, Apps, Results, Sharing and Collaboration.”

Pienaar argued that even today, most medical clinical applications follow 1990’s vintage interfaces and patient medical data is housed in locked-down silos that don’t facilitate cross-comparison. From an integration perspective, medical data is dead on arrival – the only common central point to all medical data in a hospital is the billing department. It is largely impossible for physicians or researchers to make meaningful comparisons of data housed in different silos. The system was designed to facilitate billing, and to protect patient privacy, but it impedes research and diagnostics. Pienaar believes that the current infrastructure is not at all well suited to handle the amount and complexity of data that is coming down the pipe in the next five years.

That’s when ChRIS was born. The Boston Children’s Hospital Research Integration System (ChRIS) is a novel browser-based data storage and data processing workflow manager. Developed using Javascript and Web2.0, ChRIS anonymizes personal information, and includes a plug-in to accommodate clusters of collaborators that could, potentially, be located around the world. By employing a local virtual machine on the desktop, ChRIS hides the complexity of data scheduling on an HPC by keeping everything local. The system also uses Google real-time API services that facilitate to enable real-time true image sharing and collaboration. Multiple parties can interact with the same medical visual scene, each in their own browsers, and all views, slices, are updated in real time – using a design approach that is not screen sharing, but more akin to how multi-player 3D gaming works.

SADC HPC Forum

SADC Forum 2015The SADC HPC Forum has been meeting since 2013 when the University of Texas, U.S. donated a decommissioned HPC system called “Ranger” to CHPC. More than 20 Ranger racks were divided into several mini-Rangers that were placed in eight centers in five SADC states. Each mini-Ranger forms a footprint for human capital development (HCD) and the initiative has launched an important public-private dialogue that SADC hopes will inspire support for future expansion.

Tshiamo Motshegwa (U-Botswana, SADC Forum Chair)
Tshiamo Motshegwa (U-Botswana, SADC Forum Chair)

Ms. Mmampei Chaba (Dept. of Science & Technology, SA) welcomed everyone on behalf of the host country, and Ms. Anneline Morgan (SADC Secretariat) introduced Forum Chair Tshiamo Motshegwa (U-Botswana).

Forum delegates and international advisers further refined the SADC collaborative framework document that last year’s forum attendees had begun to draft, and a final version will be presented to the SADC Ministry. There were several newcomers; delegates who hadn’t attended in the past, and representatives from countries that were entirely new to the forum, including Mauritius, Namibia and Seychelles. All were eager to know how their countries could contribute to and engage with the shared CI.

Everyone was eager to discuss best practices relating to data sharing across national borders; an uncomfortable concept for those who haven’t managed research data. The international advisers explained that the ability to facilitate the secure, seamless and reliable transfer of data among SADC sites, and beyond, is crucial to the project’s success. To that end, following the meeting, Motshegwa began to plan a cybersecurity and data conference scheduled for mid-April, 2016 in Gaborone, Botswana. U.S. and European cybersecurity experts are invited to participate. They will share roughly 50 years of collective experience with the facilitation of secure data across peered networks, international computational systems and interfederated CIs.

SADC HPC Forum delegate questions frame future discussions and HCD programming. While technology training has been the focus, each site is encouraged to add education, outreach, communication, and external relations skills to their teams. By building teams that include a combination of technical and soft skills, they will be better prepared to support and sustain their CI programs for the future.

The CHPC National Meeting, co-located meetings and student programs would not have been possible without the generosity of sponsors IntelDellAltairSeagateEclipse HoldingsMellanox, and Bright Computing. We appreciate their continued support,” said Conference Program Planner Happy Sithole (CHPC).

Presentations and videos from the awards evening are available on the CHPC website.

2016 CHPC National Meeting

Please join us in East London, South Africa for the 2016 CHPC National Meeting, and SADC HPC Forum Dec. 5-9, 2016.

Photo by Adel Groenewald
Photo by Adel Groenewald

Plan to spend extra time while visiting this beautiful region, and be sure to pack your hiking boots! It will be summer in South Africa, and you’ll find some of the best hiking, bird watching, fishing, swimming, and horseback riding in the world. East London is South Africa’s only river port city. The nearby Wild Coast region has miles of unspoiled, white beaches framed by thick forests that give way to steep, rocky cliffs. The area is steeped in Xhosa tradition, and their farms are scattered along the coast. It’s sparsely-populated, and it’s likely you’ll meet more cows than people. With a favorable international currency exchange rate, you’ll only be limited your energy and time.

 

About the Author

ElizabethLeake-headshotElizabeth Leake is the president and founder of STEM-Trek Nonprofit, a global, grassroots organization that supports scholarly travel for science, technology, engineering, and mathematics scholars from underrepresented groups and regions. Since 2011, she has worked as a communications consultant for a variety of education and research organizations, and served as correspondent for activities sponsored by the eXtreme Science and Engineering Discovery Environment (TeraGrid/XSEDE), the Partnerhip for Advanced Computing in Europe (DEISA/PRACE), the European Grid Infrastructure (EGEE/EGI), South African Center for High Performance Computing (CHPC), the Southern African Development Community (SADC), and Sustainable Horizons, Inc.

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industy updates delivered to you every week!

Fluid HPC: How Extreme-Scale Computing Should Respond to Meltdown and Spectre

February 15, 2018

The Meltdown and Spectre vulnerabilities are proving difficult to fix, and initial experiments suggest security patches will cause significant performance penalties to HPC applications. Even as these patches are rolled o Read more…

By Pete Beckman

Intel Touts Silicon Spin Qubits for Quantum Computing

February 14, 2018

Debate around what makes a good qubit and how best to manufacture them is a sprawling topic. There are many insistent voices favoring one or another approach. Referencing a paper published today in Nature, Intel has offe Read more…

By John Russell

Brookhaven Ramps Up Computing for National Security Effort

February 14, 2018

Last week, Dan Coats, the director of Director of National Intelligence for the U.S., warned the Senate Intelligence Committee that Russia was likely to meddle in the 2018 mid-term U.S. elections, much as it stands accused of doing in the 2016 Presidential election. Read more…

By John Russell

HPE Extreme Performance Solutions

Safeguard Your HPC Environment with the World’s Most Secure Industry Standard Servers

Today’s organizations operate in an environment with ever-evolving threats, and in order to protect themselves they must continuously bolster their security strategy. Hewlett Packard Enterprise (HPE) and Intel® are addressing modern security challenges with the world’s most secure industry standard servers powered by the latest generation of Intel® Xeon® Scalable processors. Read more…

AI Cloud Competition Heats Up: Google’s TPUs, Amazon Building AI Chip

February 12, 2018

Competition in the white hot AI (and public cloud) market pits Google against Amazon this week, with Google offering AI hardware on its cloud platform intended to make it easier, faster and cheaper to train and run machi Read more…

By Doug Black

Fluid HPC: How Extreme-Scale Computing Should Respond to Meltdown and Spectre

February 15, 2018

The Meltdown and Spectre vulnerabilities are proving difficult to fix, and initial experiments suggest security patches will cause significant performance penal Read more…

By Pete Beckman

Brookhaven Ramps Up Computing for National Security Effort

February 14, 2018

Last week, Dan Coats, the director of Director of National Intelligence for the U.S., warned the Senate Intelligence Committee that Russia was likely to meddle in the 2018 mid-term U.S. elections, much as it stands accused of doing in the 2016 Presidential election. Read more…

By John Russell

AI Cloud Competition Heats Up: Google’s TPUs, Amazon Building AI Chip

February 12, 2018

Competition in the white hot AI (and public cloud) market pits Google against Amazon this week, with Google offering AI hardware on its cloud platform intended Read more…

By Doug Black

Russian Nuclear Engineers Caught Cryptomining on Lab Supercomputer

February 12, 2018

Nuclear scientists working at the All-Russian Research Institute of Experimental Physics (RFNC-VNIIEF) have been arrested for using lab supercomputing resources to mine crypto-currency, according to a report in Russia’s Interfax News Agency. Read more…

By Tiffany Trader

The Food Industry’s Next Journey — from Mars to Exascale

February 12, 2018

Global food producer and one of the world's leading chocolate companies Mars Inc. has a unique perspective on the impact that exascale computing will have on the food industry. Read more…

By Scott Gibson, Oak Ridge National Laboratory

Singularity HPC Container Start-Up – Sylabs – Emerges from Stealth

February 8, 2018

The driving force behind Singularity, the popular HPC container technology, is bringing the open source platform to the enterprise with the launch of a new vent Read more…

By George Leopold

Dell EMC Debuts PowerEdge Servers with AMD EPYC Chips

February 6, 2018

AMD notched another EPYC processor win today with Dell EMC’s introduction of three PowerEdge servers (R6415, R7415, and R7425) based on the EPYC 7000-series p Read more…

By John Russell

‘Next Generation’ Universe Simulation Is Most Advanced Yet

February 5, 2018

The research group that gave us the most detailed time-lapse simulation of the universe’s evolution in 2014, spanning 13.8 billion years of cosmic evolution, is back in the spotlight with an even more advanced cosmological model that is providing new insights into how black holes influence the distribution of dark matter, how heavy elements are produced and distributed, and where magnetic fields originate. Read more…

By Tiffany Trader

Inventor Claims to Have Solved Floating Point Error Problem

January 17, 2018

"The decades-old floating point error problem has been solved," proclaims a press release from inventor Alan Jorgensen. The computer scientist has filed for and Read more…

By Tiffany Trader

Japan Unveils Quantum Neural Network

November 22, 2017

The U.S. and China are leading the race toward productive quantum computing, but it's early enough that ultimate leadership is still something of an open questi Read more…

By Tiffany Trader

AMD Showcases Growing Portfolio of EPYC and Radeon-based Systems at SC17

November 13, 2017

AMD’s charge back into HPC and the datacenter is on full display at SC17. Having launched the EPYC processor line in June along with its MI25 GPU the focus he Read more…

By John Russell

Researchers Measure Impact of ‘Meltdown’ and ‘Spectre’ Patches on HPC Workloads

January 17, 2018

Computer scientists from the Center for Computational Research, State University of New York (SUNY), University at Buffalo have examined the effect of Meltdown Read more…

By Tiffany Trader

Nvidia Responds to Google TPU Benchmarking

April 10, 2017

Nvidia highlights strengths of its newest GPU silicon in response to Google's report on the performance and energy advantages of its custom tensor processor. Read more…

By Tiffany Trader

IBM Begins Power9 Rollout with Backing from DOE, Google

December 6, 2017

After over a year of buildup, IBM is unveiling its first Power9 system based on the same architecture as the Department of Energy CORAL supercomputers, Summit a Read more…

By Tiffany Trader

Fast Forward: Five HPC Predictions for 2018

December 21, 2017

What’s on your list of high (and low) lights for 2017? Volta 100’s arrival on the heels of the P100? Appearance, albeit late in the year, of IBM’s Power9? Read more…

By John Russell

Russian Nuclear Engineers Caught Cryptomining on Lab Supercomputer

February 12, 2018

Nuclear scientists working at the All-Russian Research Institute of Experimental Physics (RFNC-VNIIEF) have been arrested for using lab supercomputing resources to mine crypto-currency, according to a report in Russia’s Interfax News Agency. Read more…

By Tiffany Trader

Leading Solution Providers

Chip Flaws ‘Meltdown’ and ‘Spectre’ Loom Large

January 4, 2018

The HPC and wider tech community have been abuzz this week over the discovery of critical design flaws that impact virtually all contemporary microprocessors. T Read more…

By Tiffany Trader

Perspective: What Really Happened at SC17?

November 22, 2017

SC is over. Now comes the myriad of follow-ups. Inboxes are filled with templated emails from vendors and other exhibitors hoping to win a place in the post-SC thinking of booth visitors. Attendees of tutorials, workshops and other technical sessions will be inundated with requests for feedback. Read more…

By Andrew Jones

How Meltdown and Spectre Patches Will Affect HPC Workloads

January 10, 2018

There have been claims that the fixes for the Meltdown and Spectre security vulnerabilities, named the KPTI (aka KAISER) patches, are going to affect applicatio Read more…

By Rosemary Francis

GlobalFoundries, Ayar Labs Team Up to Commercialize Optical I/O

December 4, 2017

GlobalFoundries (GF) and Ayar Labs, a startup focused on using light, instead of electricity, to transfer data between chips, today announced they've entered in Read more…

By Tiffany Trader

Tensors Come of Age: Why the AI Revolution Will Help HPC

November 13, 2017

Thirty years ago, parallel computing was coming of age. A bitter battle began between stalwart vector computing supporters and advocates of various approaches to parallel computing. IBM skeptic Alan Karp, reacting to announcements of nCUBE’s 1024-microprocessor system and Thinking Machines’ 65,536-element array, made a public $100 wager that no one could get a parallel speedup of over 200 on real HPC workloads. Read more…

By John Gustafson & Lenore Mullin

Flipping the Flops and Reading the Top500 Tea Leaves

November 13, 2017

The 50th edition of the Top500 list, the biannual publication of the world’s fastest supercomputers based on public Linpack benchmarking results, was released Read more…

By Tiffany Trader

V100 Good but not Great on Select Deep Learning Aps, Says Xcelerit

November 27, 2017

Wringing optimum performance from hardware to accelerate deep learning applications is a challenge that often depends on the specific application in use. A benc Read more…

By John Russell

2017 Gordon Bell Prize Finalists Named

October 23, 2017

The three finalists for this year’s Gordon Bell Prize in High Performance Computing have been announced. They include two papers on projects run on China’s Read more…

By John Russell

  • arrow
  • Click Here for More Headlines
  • arrow
Share This