Feeds

Database gurus slammed for Google post

MapReduce: a "major step backwards"

Remote control for virtualized desktops

A database pioneer and honored computer science professor have come under heavy fire for issuing a strong critique of Google's MapReduce technology for processing large unstructured databases.

Ingres inventor and Postgres architect Mike Stonebraker and his colleague, University of Wisconsin computer science professor David DeWitt, have been accused of "not getting" data in the clouds while others have demanded the duo retract what's been branded a "highly inaccurate article".

Stonebraker and DeWitt had criticized MapReduce and slammed moves to introduce MapReduce into the academic curriculum.

They called MapReduce a major step backwards because it is "sub optimal", lacks the features commonly associated with database management systems (DBMS) and is incompatible with "all of the tools DBMS users have come to depend on". They also said that it is not '"novel". They conclude that MapReduce ignores many of the developments in parallel DBMS technology over the last 25 years.

Their joint blog post drew fire from bloggers and a barrage of commentators coming out in support of MapReduce, including a detailed riposte that claimed DeWitt and Stonebraker don't know what they are talking about.

The gist of the counter argument is that MapReduce can't be compared to a relational DBMS because it is a technique for dealing with large amounts of unstructured data rather than the formal tabular data in relational DBMS. Google reckons it processes 20PB of unstructured data a day using MapReduce.

The almost complete lack of support for the view put forward by DeWitt and Stonebraker suggests they might well have misunderstood MapReduce's role in modern data processing.

Given the eminent background of both academics, though, this is surprising. DeWitt has researched large parallel DBMS since the 1980s and, in addition to his pioneering work on Ingres and Postgres, Stonebraker is currently active in the large DBMS area with his new company Vertica.

DeWitt has published more than 100 technical papers and been honored for contributions to database systems having started in the mid 1970s on a NASA- and DARPA-funded project looking at scalable object-relational system for managing very large geo-spatial data sets.

Requests for an interview have yet to be answered.®

Beginner's guide to SSL certificates

More from The Register

next story
Download alert: Nearly ALL top 100 Android, iOS paid apps hacked
Attack of the Clones? Yeah, but much, much scarier – report
Euro Parliament VOTES to BREAK UP GOOGLE. Er, OK then
It CANNA do it, captain.They DON'T have the POWER!
NSA SOURCE CODE LEAK: Information slurp tools to appear online
Now you can run your own intelligence agency
Post-Microsoft, post-PC programming: The portable REVOLUTION
Code jockeys: count up and grab your fabulous tablets
Microsoft: Your Linux Docker containers are now OURS to command
New tool lets admins wrangle Linux apps from Windows
prev story

Whitepapers

Free virtual appliance for wire data analytics
The ExtraHop Discovery Edition is a free virtual appliance will help you to discover the performance of your applications across the network, web, VDI, database, and storage tiers.
A strategic approach to identity relationship management
ForgeRock commissioned Forrester to evaluate companies’ IAM practices and requirements when it comes to customer-facing scenarios versus employee-facing ones.
10 threats to successful enterprise endpoint backup
10 threats to a successful backup including issues with BYOD, slow backups and ineffective security.
High Performance for All
While HPC is not new, it has traditionally been seen as a specialist area – is it now geared up to meet more mainstream requirements?
Website security in corporate America
Find out how you rank among other IT managers testing your website's vulnerabilities.