Channels ▼
RSS

Parallel

Better Computer "Vision"


Taking inspiration from genetic screening techniques, researchers from MIT and Harvard have demonstrated a way to build better artificial visual systems with the help of low-cost, high-performance gaming hardware.

The neural processing involved in visually recognizing even the simplest object in a natural environment is profound -- and profoundly difficult to mimic. Neuroscientists have made broad advances in understanding the visual system, but much of the inner workings of biologically based systems remain a mystery.

Using Graphics Processing Units (GPUs) -- the same technology video game designers use to render life-like graphics -- MIT and Harvard researchers are now making progress faster than ever before. "We made a powerful computing system that delivers over hundred fold speed-ups relative to conventional methods," said Nicolas Pinto, a PhD candidate in James DiCarlo's lab at the McGovern Institute for Brain Research at MIT. "With this extra computational power, we can discover new vision models that traditional methods miss." Pinto co-authored the PLoS Computational Biology study, titled A High-Throughput Screening Approach to Discovering Good Forms of Biologically Inspired Visual Representation with David Cox of the Visual Neuroscience Group at the Rowland Institute at Harvard.

Harnessing the processing power of dozens of high-performance NVIDIA graphics cards and PlayStation 3s gaming devices, the team designed a high-throughput screening process to tease out the best parameters for visual object recognition tasks. The resulting model outperformed a crop of state-of-the-art vision systems across a range of tests -- more accurately identifying a range of objects on random natural backgrounds with variation in position, scale, and rotation. Had the team used conventional computational tools, the one-week screening phase would have taken over two years to complete.

The researchers say that their high-throughput approach could be applied to other areas of computer vision, such as face identification, object tracking, pedestrian detection for automotive applications, and gesture and action recognition. Moreover, as scientists better understand what components make a good artificial vision system, they can use these hints to better understand the human brain as well.

The study was funded by the National Institutes of Health, McKnight Endowment for Neuroscience, Jerry and Marge Burnett, the McGovern Institute for Brain Research at MIT, and the Rowland Institute at Harvard. Hardware support was provided by the NVIDIA Corporation.


Related Reading


More Insights






Currently we allow the following HTML tags in comments:

Single tags

These tags can be used alone and don't need an ending tag.

<br> Defines a single line break

<hr> Defines a horizontal line

Matching tags

These require an ending tag - e.g. <i>italic text</i>

<a> Defines an anchor

<b> Defines bold text

<big> Defines big text

<blockquote> Defines a long quotation

<caption> Defines a table caption

<cite> Defines a citation

<code> Defines computer code text

<em> Defines emphasized text

<fieldset> Defines a border around elements in a form

<h1> This is heading 1

<h2> This is heading 2

<h3> This is heading 3

<h4> This is heading 4

<h5> This is heading 5

<h6> This is heading 6

<i> Defines italic text

<p> Defines a paragraph

<pre> Defines preformatted text

<q> Defines a short quotation

<samp> Defines sample computer code text

<small> Defines small text

<span> Defines a section in a document

<s> Defines strikethrough text

<strike> Defines strikethrough text

<strong> Defines strong text

<sub> Defines subscripted text

<sup> Defines superscripted text

<u> Defines underlined text

Dr. Dobb's encourages readers to engage in spirited, healthy debate, including taking us to task. However, Dr. Dobb's moderates all comments posted to our site, and reserves the right to modify or remove any content that it determines to be derogatory, offensive, inflammatory, vulgar, irrelevant/off-topic, racist or obvious marketing or spam. Dr. Dobb's further reserves the right to disable the profile of any commenter participating in said activities.

 
Disqus Tips To upload an avatar photo, first complete your Disqus profile. | View the list of supported HTML tags you can use to style comments. | Please read our commenting policy.
 

Video