Feeds

Database gurus slammed for Google post

MapReduce: a "major step backwards"

New hybrid storage solutions

A database pioneer and honored computer science professor have come under heavy fire for issuing a strong critique of Google's MapReduce technology for processing large unstructured databases.

Ingres inventor and Postgres architect Mike Stonebraker and his colleague, University of Wisconsin computer science professor David DeWitt, have been accused of "not getting" data in the clouds while others have demanded the duo retract what's been branded a "highly inaccurate article".

Stonebraker and DeWitt had criticized MapReduce and slammed moves to introduce MapReduce into the academic curriculum.

They called MapReduce a major step backwards because it is "sub optimal", lacks the features commonly associated with database management systems (DBMS) and is incompatible with "all of the tools DBMS users have come to depend on". They also said that it is not '"novel". They conclude that MapReduce ignores many of the developments in parallel DBMS technology over the last 25 years.

Their joint blog post drew fire from bloggers and a barrage of commentators coming out in support of MapReduce, including a detailed riposte that claimed DeWitt and Stonebraker don't know what they are talking about.

The gist of the counter argument is that MapReduce can't be compared to a relational DBMS because it is a technique for dealing with large amounts of unstructured data rather than the formal tabular data in relational DBMS. Google reckons it processes 20PB of unstructured data a day using MapReduce.

The almost complete lack of support for the view put forward by DeWitt and Stonebraker suggests they might well have misunderstood MapReduce's role in modern data processing.

Given the eminent background of both academics, though, this is surprising. DeWitt has researched large parallel DBMS since the 1980s and, in addition to his pioneering work on Ingres and Postgres, Stonebraker is currently active in the large DBMS area with his new company Vertica.

DeWitt has published more than 100 technical papers and been honored for contributions to database systems having started in the mid 1970s on a NASA- and DARPA-funded project looking at scalable object-relational system for managing very large geo-spatial data sets.

Requests for an interview have yet to be answered.®

Secure remote control for conventional and virtual desktops

More from The Register

next story
'Windows 9' LEAK: Microsoft's playing catchup with Linux
Multiple desktops and live tiles in restored Start button star in new vids
Not appy with your Chromebook? Well now it can run Android apps
Google offers beta of tricky OS-inside-OS tech
New 'Cosmos' browser surfs the net by TXT alone
No data plan? No WiFi? No worries ... except sluggish download speed
Greater dev access to iOS 8 will put us AT RISK from HACKERS
Knocking holes in Apple's walled garden could backfire, says securo-chap
NHS grows a NoSQL backbone and rips out its Oracle Spine
Open source? In the government? Ha ha! What, wait ...?
Google extends app refund window to two hours
You now have 120 minutes to finish that game instead of 15
Intel: Hey, enterprises, drop everything and DO HADOOP
Big Data analytics projected to run on more servers than any other app
prev story

Whitepapers

Secure remote control for conventional and virtual desktops
Balancing user privacy and privileged access, in accordance with compliance frameworks and legislation. Evaluating any potential remote control choice.
Saudi Petroleum chooses Tegile storage solution
A storage solution that addresses company growth and performance for business-critical applications of caseware archive and search along with other key operational systems.
High Performance for All
While HPC is not new, it has traditionally been seen as a specialist area – is it now geared up to meet more mainstream requirements?
Security for virtualized datacentres
Legacy security solutions are inefficient due to the architectural differences between physical and virtual environments.
Providing a secure and efficient Helpdesk
A single remote control platform for user support is be key to providing an efficient helpdesk. Retain full control over the way in which screen and keystroke data is transmitted.