Feeds

Amazon teaches cloud to speak Pig Latin

Adoophay orfay ethay assesmay

Next gen security for virtualised datacentres

Amazon has taught its cloud to speak Pig Latin.

In April, Jeff Bezos and company unveiled a new web service based on Hadoop, the open-source phenomenon that seeks to mimic MapReduce, the distributed data-crunching platform that drives Google's online infrastructure. And today, Amazon announced that its Elastic MapReduce service now includes support for Pig Latin, the Hadoop programming language first developed at Yahoo!.

You can write straight to Hadoop in Java, but Pig puts Hadoop programming on a somewhat higher level. As Amazon puts it: "Pig Latin is a SQL-like data transformation language. You can use Pig Latin to run complex processes on large-scale compute clusters without having to spend time learning the MapReduce paradigm."

But the SQL comparison is a tad misleading. Hive - a Hadoop programming language seeded by Facebook - is closer to SQL, as is a second as-yet-unnamed language under development at Yahoo!. Pig sits somewhere between the SQL-like paradigm and the low-level code of MapReduce.

Amazon offers two means of using Pig: an "Interactive mode," which lets you run Pig queries on an existing MapReduce cluster by setting up an secure shell connection, and a "batch mode," which involves launching multiple MapReduce server instances that reference your Pig Latin.

Based on Google-published research papers, Hadoop mimics the company's MapReduce framework, which maps data-crunching tasks across distributed machines, splitting them into sub-tasks, before reducing the results into one master calculation. Thus Amazon Elastic MapReduce.

The Apache-hosted Hadoop was originally developed by Nutch-crawler founder Doug Cutting. After three and half years at Yahoo! developing the platform, Cutting is now headed for Cloudera, a Silicon Valley startup that has commercialized Hadoop - Red Hat-style. ®

Build a business case: developing custom apps

More from The Register

next story
Why has the web gone to hell? Market chaos and HUMAN NATURE
Tim Berners-Lee isn't happy, but we should be
Mozilla's 'Tiles' ads debut in new Firefox nightlies
You can try turning them off and on again
Microsoft boots 1,500 dodgy apps from the Windows Store
DEVELOPERS! DEVELOPERS! DEVELOPERS! Naughty, misleading developers!
'Stop dissing Google or quit': OK, I quit, says Code Club co-founder
And now a message from our sponsors: 'STFU or else'
Apple promises to lift Curse of the Drained iPhone 5 Battery
Have you tried turning it off and...? Never mind, here's a replacement
Uber, Lyft and cutting corners: The true face of the Sharing Economy
Casual labour and tired ideas = not really web-tastic
Linux turns 23 and Linus Torvalds celebrates as only he can
No, not with swearing, but by controlling the release cycle
prev story

Whitepapers

Gartner critical capabilities for enterprise endpoint backup
Learn why inSync received the highest overall rating from Druva and is the top choice for the mobile workforce.
Implementing global e-invoicing with guaranteed legal certainty
Explaining the role local tax compliance plays in successful supply chain management and e-business and how leading global brands are addressing this.
Rethinking backup and recovery in the modern data center
Combining intelligence, operational analytics, and automation to enable efficient, data-driven IT organizations using the HP ABR approach.
Consolidation: The Foundation for IT Business Transformation
In this whitepaper learn how effective consolidation of IT and business resources can enable multiple, meaningful business benefits.
Next gen security for virtualised datacentres
Legacy security solutions are inefficient due to the architectural differences between physical and virtual environments.