Feeds

The case for open source ETL

What you see is what you get

Choosing a cloud hosting partner with confidence

Comment As far as I have been able to discover, there are four open source ETL (extract, transform and load) tools on the market. Somewhat surprisingly, two of them are homonyms: KETL and Kettle, the other two being Enhydra Octopus and CloverETL.

Kettle is based on an ETTL paradigm, the extra ‘T’ standing for transport (which seems an unnecessary complication) and wins the prize for the product with the most sense of humour as it has four components that are variously named Spoon, Pan, Chef and Kitchen.

The most interesting question is where the market for open source ETL is. Looking at the products one would have to assume that they are mostly in the same space as the Sesame Software product that I discussed recently. That is, they are aimed at developers that know what they are doing and do not need (and do not get) a graphical drag-and-drop style product. The exception is Kettle, which looks much more like an PowerCenter or DataStage.

Another difference in open source products can be in implementation. Kinetic Networks, for example, the developers of KETL, reckons that you may need some implementation assistance with its product. In part, this is a result of the product’s origins: it was originally developed for in-house use and in conjunction with professional services engagements, so it is not really surprising if there are aspects of the product that have not been automated for the open source market yet.

In general, what you see is what you get with open source products, though there are add-on products for both of the homonymous products. In the case of Kettle one of the partners offers an SAP connector while Kinetic Networks has a number of options that it offers in conjunction with KETL, notably an MPP (massively parallel processing) option for improved performance, a data profiling extension and a clickstream capability.

As far as I can see there are no options (apart from support) available with either Enhydra Octopus or CloverETL, the latter being a product that generates Java. This, despite its attractions (especially for ISVs), is still a relatively rare capability: ETL Solutions has a product that generates Java and ETI plans to, but otherwise this is not generally available, so it represents a potential market for CloverETL that is not available to its open source counterparts.

Enhydra Octopus is distinguished by the fact that it has different companies offering support for the product in Europe, Japan and the United States whereas the other products only have limited support options (USA for KETL, Austria/Belgium for Kettle and the Czech Republic for Enhydra Octopus).

In other words, each of the products has something different going for it, though none of them will trouble the likes of Informatica, IBM or Ab Initio.

Copyright © 2005, IT-Analysis.com

Intelligent flash storage arrays

More from The Register

next story
Netscape Navigator - the browser that started it all - turns 20
It was 20 years ago today, Marc Andreeesen taught the band to play
Sway: Microsoft's new Office app doesn't have an Undo function
Content aggregation, meet the workplace ... oh
Sign off my IT project or I’ll PHONE your MUM
Honestly, it’s a piece of piss
Return of the Jedi – Apache reclaims web server crown
.london, .hamburg and .公司 - that's .com in Chinese - storm the web server charts
NetWare sales revive in China thanks to that man Snowden
If it ain't Microsoft, it's in fashion behind the Great Firewall
Chrome 38's new HTML tag support makes fatties FIT and SKINNIER
First browser to protect networks' bandwith using official spec
Admins! Never mind POODLE, there're NEW OpenSSL bugs to splat
Four new patches for open-source crypto libraries
prev story

Whitepapers

Forging a new future with identity relationship management
Learn about ForgeRock's next generation IRM platform and how it is designed to empower CEOS's and enterprises to engage with consumers.
Cloud and hybrid-cloud data protection for VMware
Learn how quick and easy it is to configure backups and perform restores for VMware environments.
Three 1TB solid state scorchers up for grabs
Big SSDs can be expensive but think big and think free because you could be the lucky winner of one of three 1TB Samsung SSD 840 EVO drives that we’re giving away worth over £300 apiece.
Reg Reader Research: SaaS based Email and Office Productivity Tools
Read this Reg reader report which provides advice and guidance for SMBs towards the use of SaaS based email and Office productivity tools.
Security for virtualized datacentres
Legacy security solutions are inefficient due to the architectural differences between physical and virtual environments.