Feeds

Google trumpets PageRank for pics

Image search gets real

Boost IT visibility and business value

Nearly a decade ago, Google unveiled an algorithm called PageRank, reinventing the way we search for web pages. Now, the company says, it has a technology that can do much the same for online image search.

Last week, at the International World Wide Web Conference in Beijing, two Google-affiliated researchers presented a paper called "PageRank for Product Image Search," trumpeting a fledging algorithm that overhauls the primitive text-based methods used by the company's current image search technologies.

"Our experiment results show significant improvement, in terms of user satisfaction and relevancy, in comparison to the most recent Google Image Search results," Shumeet Baluja and Yushi Jing tell the world from the pages of their research paper, available here.

Of course, the most recent Google Image Search results are often rubbish. Currently, when ranking images, the big search engines spend little time examining the images themselves. Instead, they look at the text surrounding those images.

By contrast, Google's PageRank for Product Image Search - also known as "VisualRank" - seeks to actually understand what's pictured. But the technology goes beyond classic image recognition, which can be time consuming and/or expensive - and which often breaks down with anything other than faces and a handful of other image types. In an effort to properly identify a wider range of objects, Baluja and Jing have merged existing image processing techniques with the sort of "link analysis" made famous by PageRank.

"Through an iterative procedure based on the PageRank computation, a numerical weight is assigned to each image," they explain. "This measures its relative importance to the other images being considered."

With classic image recognition, you typically take a known image and compare it to other images. You might use a known photo of Paris Hilton, for instance, to find other Paris pics. But VisualRank takes a different tack. Google's algorithm looks for "visual themes" across a collection of images, before ranking each image based on how well it matches those themes.

As an example, the researchers point to an image search on the word "McDonald's." In this case, VisualRank might identify the famous golden arches as theme. An image dominated by the golden arches would then be ranked higher than a pic where the arches are tucked into the background.

Baluja and Jing recently tested their algorithm using images retrieved by Google's 2000 most popular product searches, and a panel of 150 people decided that VisualRank reduced the number of irrelevant results by 83 per cent. The question is whether this could be applied to Google's entire database of images.

At the moment, this is just a research paper. And Google isn't the first to toy with the idea of true image search. After launching an online photo sharing tool that included face and character recognition, the Silicon Valley based Riya is now offering an image-rec shopping engine, known as Like.com, that locates products on sale across the web. And the transatlantic image rec gurus at Blinkx are well on their way with video search. ®

Boost IT visibility and business value

More from The Register

next story
The Return of BSOD: Does ANYONE trust Microsoft patches?
Sysadmins, you're either fighting fires or seen as incompetents now
Munich considers dumping Linux for ... GULP ... Windows!
Give a penguinista a hug, the Outlook's not good for open source's poster child
Intel's Raspberry Pi rival Galileo can now run Windows
Behold the Internet of Things. Wintel Things
Linux Foundation says many Linux admins and engineers are certifiable
Floats exam program to help IT employers lock up talent
Microsoft cries UNINSTALL in the wake of Blue Screens of Death™
Cache crash causes contained choloric calamity
Eat up Martha! Microsoft slings handwriting recog into OneNote on Android
Freehand input on non-Windows kit for the first time
prev story

Whitepapers

Implementing global e-invoicing with guaranteed legal certainty
Explaining the role local tax compliance plays in successful supply chain management and e-business and how leading global brands are addressing this.
Top 10 endpoint backup mistakes
Avoid the ten endpoint backup mistakes to ensure that your critical corporate data is protected and end user productivity is improved.
Top 8 considerations to enable and simplify mobility
In this whitepaper learn how to successfully add mobile capabilities simply and cost effectively.
Rethinking backup and recovery in the modern data center
Combining intelligence, operational analytics, and automation to enable efficient, data-driven IT organizations using the HP ABR approach.
Reg Reader Research: SaaS based Email and Office Productivity Tools
Read this Reg reader report which provides advice and guidance for SMBs towards the use of SaaS based email and Office productivity tools.