Feeds

Amazon teaches cloud to speak Pig Latin

Adoophay orfay ethay assesmay

The essential guide to IT transformation

Amazon has taught its cloud to speak Pig Latin.

In April, Jeff Bezos and company unveiled a new web service based on Hadoop, the open-source phenomenon that seeks to mimic MapReduce, the distributed data-crunching platform that drives Google's online infrastructure. And today, Amazon announced that its Elastic MapReduce service now includes support for Pig Latin, the Hadoop programming language first developed at Yahoo!.

You can write straight to Hadoop in Java, but Pig puts Hadoop programming on a somewhat higher level. As Amazon puts it: "Pig Latin is a SQL-like data transformation language. You can use Pig Latin to run complex processes on large-scale compute clusters without having to spend time learning the MapReduce paradigm."

But the SQL comparison is a tad misleading. Hive - a Hadoop programming language seeded by Facebook - is closer to SQL, as is a second as-yet-unnamed language under development at Yahoo!. Pig sits somewhere between the SQL-like paradigm and the low-level code of MapReduce.

Amazon offers two means of using Pig: an "Interactive mode," which lets you run Pig queries on an existing MapReduce cluster by setting up an secure shell connection, and a "batch mode," which involves launching multiple MapReduce server instances that reference your Pig Latin.

Based on Google-published research papers, Hadoop mimics the company's MapReduce framework, which maps data-crunching tasks across distributed machines, splitting them into sub-tasks, before reducing the results into one master calculation. Thus Amazon Elastic MapReduce.

The Apache-hosted Hadoop was originally developed by Nutch-crawler founder Doug Cutting. After three and half years at Yahoo! developing the platform, Cutting is now headed for Cloudera, a Silicon Valley startup that has commercialized Hadoop - Red Hat-style. ®

Secure remote control for conventional and virtual desktops

More from The Register

next story
Munich considers dumping Linux for ... GULP ... Windows!
Give a penguinista a hug, the Outlook's not good for open source's poster child
The Return of BSOD: Does ANYONE trust Microsoft patches?
Sysadmins, you're either fighting fires or seen as incompetents now
Intel's Raspberry Pi rival Galileo can now run Windows
Behold the Internet of Things. Wintel Things
Microsoft cries UNINSTALL in the wake of Blue Screens of Death™
Cache crash causes contained choloric calamity
Time to move away from Windows 7 ... whoa, whoa, who said anything about Windows 8?
Start migrating now to avoid another XPocalypse – Gartner
Eat up Martha! Microsoft slings handwriting recog into OneNote on Android
Freehand input on non-Windows kit for the first time
You'll find Yoda at the back of every IT conference
The piss always taking is he. Bastard the.
prev story

Whitepapers

Endpoint data privacy in the cloud is easier than you think
Innovations in encryption and storage resolve issues of data privacy and key requirements for companies to look for in a solution.
Implementing global e-invoicing with guaranteed legal certainty
Explaining the role local tax compliance plays in successful supply chain management and e-business and how leading global brands are addressing this.
Top 8 considerations to enable and simplify mobility
In this whitepaper learn how to successfully add mobile capabilities simply and cost effectively.
Solving today's distributed Big Data backup challenges
Enable IT efficiency and allow a firm to access and reuse corporate information for competitive advantage, ultimately changing business outcomes.
Reg Reader Research: SaaS based Email and Office Productivity Tools
Read this Reg reader report which provides advice and guidance for SMBs towards the use of SaaS based email and Office productivity tools.