Feeds

Amazon teaches cloud to speak Pig Latin

Adoophay orfay ethay assesmay

Boost IT visibility and business value

Amazon has taught its cloud to speak Pig Latin.

In April, Jeff Bezos and company unveiled a new web service based on Hadoop, the open-source phenomenon that seeks to mimic MapReduce, the distributed data-crunching platform that drives Google's online infrastructure. And today, Amazon announced that its Elastic MapReduce service now includes support for Pig Latin, the Hadoop programming language first developed at Yahoo!.

You can write straight to Hadoop in Java, but Pig puts Hadoop programming on a somewhat higher level. As Amazon puts it: "Pig Latin is a SQL-like data transformation language. You can use Pig Latin to run complex processes on large-scale compute clusters without having to spend time learning the MapReduce paradigm."

But the SQL comparison is a tad misleading. Hive - a Hadoop programming language seeded by Facebook - is closer to SQL, as is a second as-yet-unnamed language under development at Yahoo!. Pig sits somewhere between the SQL-like paradigm and the low-level code of MapReduce.

Amazon offers two means of using Pig: an "Interactive mode," which lets you run Pig queries on an existing MapReduce cluster by setting up an secure shell connection, and a "batch mode," which involves launching multiple MapReduce server instances that reference your Pig Latin.

Based on Google-published research papers, Hadoop mimics the company's MapReduce framework, which maps data-crunching tasks across distributed machines, splitting them into sub-tasks, before reducing the results into one master calculation. Thus Amazon Elastic MapReduce.

The Apache-hosted Hadoop was originally developed by Nutch-crawler founder Doug Cutting. After three and half years at Yahoo! developing the platform, Cutting is now headed for Cloudera, a Silicon Valley startup that has commercialized Hadoop - Red Hat-style. ®

Build a business case: developing custom apps

More from The Register

next story
KDE releases ice-cream coloured Plasma 5 just in time for summer
Melty but refreshing - popular rival to Mint's Cinnamon's still a work in progress
Leaked Windows Phone 8.1 Update specs tease details of Nokia's next mobes
New screen sizes, dual SIMs, voice over LTE, and more
Mozilla keeps its Beard, hopes anti-gay marriage troubles are now over
Plenty on new CEO's todo list – starting with Firefox's slipping grasp
Apple: We'll unleash OS X Yosemite beta on the MASSES on 24 July
Starting today, regular fanbois will be guinea pigs, it tells Reg
Another day, another Firefox: Version 31 is upon us ALREADY
Web devs, Mozilla really wants you to like this one
Secure microkernel that uses maths to be 'bug free' goes open source
Hacker-repelling, drone-protecting code will soon be yours to tweak as you see fit
Cloudy CoreOS Linux distro declares itself production-ready
Lightweight, container-happy Linux gets first Stable release
prev story

Whitepapers

Implementing global e-invoicing with guaranteed legal certainty
Explaining the role local tax compliance plays in successful supply chain management and e-business and how leading global brands are addressing this.
Boost IT visibility and business value
How building a great service catalog relieves pressure points and demonstrates the value of IT service management.
Why and how to choose the right cloud vendor
The benefits of cloud-based storage in your processes. Eliminate onsite, disk-based backup and archiving in favor of cloud-based data protection.
The Essential Guide to IT Transformation
ServiceNow discusses three IT transformations that can help CIO's automate IT services to transform IT and the enterprise.
Maximize storage efficiency across the enterprise
The HP StoreOnce backup solution offers highly flexible, centrally managed, and highly efficient data protection for any enterprise.