Visit additional Tabor Communication Publications
June 28, 2012
Search giant Google along with researchers from Stanford University have made an interesting discovery based on an X labs project. After being fed 10 million images from YouTube, a 16,000-core cluster learned how to recognize various objects, including cats. Earlier this week, the New York Times detailed the program, explaining its methods and potential use cases.
The project began a few years ago when Google researchers planned to make a human brain simulation. The compute cluster acted as the brain’s neural network, tasked with learning on its own and using the Internet as a source of information. After processing millions of unlabeled YouTube thumbnails, the system taught itself how to recognize a cat. While the video website is known for its comprehensive collection of user submitted feline antics, project researchers were focused on the simulation’s ability to learn objects without human input.
A 1,000-node cluster was the basis for the neural network, representing more than 1 billion connections. The unlabeled images were collected randomly and processed by machine-learning algorithms. Similar to IBM’s Watson, the technology relies on “deep learning” techniques, using previous outcomes to inform future decisions. Machine learning has also become integral in other applications, including speech recognition.
The researchers removed all identifying labels from the images because they wanted to see if the network was able to create the concept of an object. Dr. Jeff Dean of Stanford University explained how the project differs from other recognition technologies. “We never told it during the training, ‘This is a cat,’ ” he said “It basically invented the concept of a cat. We probably have other ones that are side views of cats.”
As a result, the software generated a vague image of a cat on its own. The simulation has also introduced possible evidence of “grandmother neurons”, which some believe are specialized cells, trained to recognize an individual face or concept.
It’s not all about cats, of course. The system was also able to identify human faces and bodies to some degree. Compared to previous attempts to identify unlabeled images, the simulation fared much better at learning and recognition. According to the researchers’ findings, they were able to deliver 15.8 percent accuracy in recognizing 20,000 object categories. They claim that’s 70 percent better than what had been achieved previously.
According to David Bader, executive director of high-performance computing at the Georgia Tech College of Computing, the simulation represents an improvement of “an order of magnitude over previous efforts.” He believes this work could lead to a complete model of the human visual cortex before the end of the decade.
Full story at The New York Times
Contributing commentator, Andrew Jones, offers a break in the news cycle with an assessment of what the national "size matters" contest means for the U.S. and other nations...
Today at the International Supercomputing Conference in Leipzing, Germany, Jack Dongarra presented on a proposed benchmark that could carry a bit more weight than its older Linpack companion. The high performance conjugate gradient (HPCG) concept takes into account new architectures for new applications, while shedding the floating point....
Not content to let the Tianhe-2 announcement ride alone, Intel rolled out a series of announcements around its Knights Corner and Xeon Phi products--all of which are aimed at adding some options and variety for a wider base of potential users across the HPC spectrum. Today at the International Supercomputing Conference, the company's Raj....
05/10/2013 | Cleversafe, Cray, DDN, NetApp, & Panasas | From Wall Street to Hollywood, drug discovery to homeland security, companies and organizations of all sizes and stripes are coming face to face with the challenges – and opportunities – afforded by Big Data. Before anyone can utilize these extraordinary data repositories, however, they must first harness and manage their data stores, and do so utilizing technologies that underscore affordability, security, and scalability.
04/15/2013 | Bull | “50% of HPC users say their largest jobs scale to 120 cores or less.” How about yours? Are your codes ready to take advantage of today’s and tomorrow’s ultra-parallel HPC systems? Download this White Paper by Analysts Intersect360 Research to see what Bull and Intel’s Center for Excellence in Parallel Programming can do for your codes.
Join HPCwire Editor Nicole Hemsoth and Dr. David Bader from Georgia Tech as they take center stage on opening night at Atlanta's first Big Data Kick Off Week, filmed in front of a live audience. Nicole and David look at the evolution of HPC, today's big data challenges, discuss real world solutions, and reveal their predictions. Exactly what does the future holds for HPC?
Join our webinar to learn how IT managers can migrate to a more resilient, flexible and scalable solution that grows with the data center. Mellanox VMS is future-proof, efficient and brings significant CAPEX and OPEX savings. The VMS is available today.