Los Alamos National Laboratory
HPCwire

Since 1986 - Covering the Fastest Computers
in the World and the People Who Run Them

Language Flags

Visit additional Tabor Communication Publications

Datanami
Digital Manufacturing Report
HPC in the Cloud
Green Computing Report

Tabor Communications
Corporate Video

NVIDIA Report: 12x CUDA Performance Gains Over CPU-Only


SANTA CLARA, Calif., Feb. 5 – NVIDIA today announced a new report describing dramatic performance increases for math libraries accelerated using NVIDIA CUDA, the world's most pervasive parallel computing platform and programming model for accelerating scientific and engineering applications on NVIDIA GPUs.

Available as a free download on the NVIDIA Developer Zone website, the report details the ground-breaking application performance increases enabled by CUDA 5 GPU-accelerated math libraries running on the new NVIDIA Tesla K20 GPU accelerators. These include cuFFT, cuBLAS, cuSPARSE, cuRAND, the NVIDIA Performance Primitives (NPP) library, Thrust and a C99-compatible math library, all of which are available for free in the latest CUDA 5 Toolkit release.

Providing application developers with the easiest and fastest way to add GPU acceleration to their code, the CUDA libraries deliver dramatic performance advantages compared to Intel's Math Kernel Library (MKL) routines running on the latest Intel CPUs. Performance highlights from the report include:

  • 12x faster sparse matrix routines(1)
  • More than 7x faster cuBLAS DGEMM performance(2)
  • Up to 16x faster tri-diagonal solver (gtsv)(3)
  • Up to 14x faster cuRAND performance(4)

NVIDIA is hosting a webinar on Wednesday, Feb. 6, from 10 – 11 a.m. PT to provide more information about CUDA 5 math library performance. Register for the webinar on the GTC Express website.

Hands-on Labs @ GTC

NVIDIA will also provide developers with a unique opportunity to experience the dramatic performance increases delivered by GPU-accelerated libraries in a series of hands-on labs at the fourth-annual GPU Technology Conference (GTC).

At GTC 2013, which will be held in San Jose, Calif., from March 18-21, developers will have the opportunity to expand their programming skills with CUDA C/C++ and OpenACC. They can also experience the performance advantages of the new CUDA 5 GPU-accelerated libraries using their own laptops to virtually access powerful GPU-based workstations hosted by Amazon in the cloud. To register for the conference, visit the GTC 2013 registration page. Once registered, attendees can log in to reserve a seat for the hands-on labs.

The CUDA 5 platform makes the development of GPU-accelerated applications faster and easier than ever, including support for dynamic parallelism, GPU-callable libraries, NVIDIA GPUDirect technology support for RDMA (remote direct memory access), and the NVIDIA Nsight Eclipse Edition integrated development environment (IDE). To learn more about CUDA or download the latest version, visit the CUDA website.

About CUDA

CUDA is a parallel computing platform and programming model developed by NVIDIA. It enables dramatic increases in computing performance by harnessing the power of GPUs. With more than 1.6 million downloads, supporting more than 180 leading engineering, scientific and commercial applications, the CUDA programming model is the most popular way for developers to take advantage of GPU-accelerated computing.

About NVIDIA

NVIDIA awakened the world to computer graphics when it invented the GPU in 1999. Today, its processors power a broad range of products from smartphones to supercomputers. NVIDIA's mobile processors are used in cell phones, tablets and auto infotainment systems. PC gamers rely on GPUs to enjoy spectacularly immersive worlds. Professionals use them to create 3D graphics and visual effects in movies and to design everything from golf clubs to jumbo jets. And researchers utilize GPUs to advance the frontiers of science with high performance computing. The company has more than 5,000 patents issued, allowed or filed, including ones covering ideas essential to modern computing. For more information, see www.nvidia.com.

-----

Source: NVIDIA Corp.

June 17, 2013

June 14, 2013

June 13, 2013

June 12, 2013

June 11, 2013

June 10, 2013

June 07, 2013

June 06, 2013

June 05, 2013


Most Read Features

Most Read Around the Web

Most Read This Just In

Asetek

Feature Articles

Intel Snaps New Grips to HPC Hook

Not content to let the Tianhe-2 announcement ride alone, Intel rolled out a series of announcements around its Knights Corner and Xeon Phi products--all of which are aimed at adding some options and variety for a wider base of potential users across the HPC spectrum. Today at the International Supercomputing Conference, the company's Raj....
Read more...

Top 500 Results Reveal Global Acceleration, Balance Shift

The Top 500 list of the world's fastest computers has just been announced. Not surprisingly, since it's been reported on prior to the official announcement, the Chinese Tianhe-2 system tops the list. And that is an understatement. We talk with Jack Dongarra, Horst Simon, Hans Meuer and others from the....
Read more...

Six Can't Miss Sessions for ISC'13

Outside of the main attractions, including the keynote sessions, vendor showdowns, Think Tank panels, BoFs, and tutorial elements, the International Supercomputing Conference has balanced its five-day agenda with some striking panels, discussions and topic areas that are worthy of some attention....
Read more...

Short Takes

Supercomputers: Still the King of the HPC Hill

Jun 17, 2013 | The advent of low-power mobile processors and cloud delivery models is changing the economics of computing. But just as an economy car is good at different things than a full size truck, an HPC workload still has certain computing demands that neither the fastest smartphone nor the most elastic cloud cluster can fulfill.
Read more...

TACC Longhorn Takes On Natural Language Processing

Jun 14, 2013 | For all the progress we've made in IT over the last 50 years, there's one area of life that has steadfastly eluded the grasp of computers: understanding human language. Now, researchers at the Texas Advanced Computing Center (TACC) are utilizing a Hadoop cluster on its Longhorn supercomputer to move the state of the art of language processing a little bit further.
Read more...

Titan Didn't Redo LINPACK for June Top 500 List

Jun 13, 2013 | Titan, the Cray XK7 at the Oak Ridge National Lab that debuted last fall as the fastest supercomputer in the world with 17.59 petaflops of sustained computing power, will rely on its previous LINPACK test for the upcoming edition of the Top 500 list.
Read more...

Top Supercomputer Signals Growth of Chinese HPC Industry

Jun 12, 2013 | At 31 petaflops of sustained LINPACK capacity, the new Chinese Tianhe-2 supercomputer will be the fastest supercomputer in the world when this month's Top 500 list comes out, as we reported previously in HPCwire.
Read more...

Intel Says Haswell Chips Offer ISVs Full OpenCL Compatibility

Jun 12, 2013 | HPC system makers are lining up to announce compatibility with the new fourth generation Intel Core processor, codenamed "Haswell." The new Iris GPUs based on the Haswell architecture are giving Intel new credibility in the graphics processing department.
Read more...

Sponsored Whitepapers

Best Practices in Big Data Storage

05/10/2013 | Cleversafe, Cray, DDN, NetApp, & Panasas | From Wall Street to Hollywood, drug discovery to homeland security, companies and organizations of all sizes and stripes are coming face to face with the challenges – and opportunities – afforded by Big Data. Before anyone can utilize these extraordinary data repositories, however, they must first harness and manage their data stores, and do so utilizing technologies that underscore affordability, security, and scalability.

Progress in Parallel: the Bull Parallel Programming Center

04/15/2013 | Bull | “50% of HPC users say their largest jobs scale to 120 cores or less.” How about yours? Are your codes ready to take advantage of today’s and tomorrow’s ultra-parallel HPC systems? Download this White Paper by Analysts Intersect360 Research to see what Bull and Intel’s Center for Excellence in Parallel Programming can do for your codes.

Sponsored Multimedia

HPCwire Live! Atlanta's Big Data Kick Off Week Meets HPC

Join HPCwire Editor Nicole Hemsoth and Dr. David Bader from Georgia Tech as they take center stage on opening night at Atlanta's first Big Data Kick Off Week, filmed in front of a live audience. Nicole and David look at the evolution of HPC, today's big data challenges, discuss real world solutions, and reveal their predictions. Exactly what does the future holds for HPC?

Webinar: Mellanox Virtual Modular Switch, the Most Efficient 40GbE Aggregation Switch Solution

Join our webinar to learn how IT managers can migrate to a more resilient, flexible and scalable solution that grows with the data center. Mellanox VMS is future-proof, efficient and brings significant CAPEX and OPEX savings. The VMS is available today.

Atlanta's Big Data Kick Off Week Meets HPC Cray Xyratex

HPC Job Bank


Featured Events






  • November 17, 2013 - November 22, 2013
    SC'13
    Denver, CO
    United States


HPCwire Events