Feeds

MongoDB daddy: My baby beats Google BigTable

The web is built on objects. Not tables

Internet Security Threat Report 2014

After a decade as chief technology officer at DoubleClick – the internet ad giant he cofounded in 1995 – Dwight Merriman set out to build a "platform cloud" along the lines of Google App Engine or Microsoft Azure. But this was before people called them platform clouds, before anyone knew about App Engine or Azure.

It was late 2007, and the idea was to create an online service for developing, hosting, and automatically scaling web applications. "What we were building was very similar to what App Engine eventually became," Merriman tells us. But unlike App Engine, the service would be underpinned by an entirely open source software stack, and somewhere along the way Merriman and his team realized that no open source database platform was suited to such a service.

"We felt like a lot of existing databases didn't really have the 'cloud computing' principles you want them to have: elasticity, scalability, and ... easy administration, but also ease of use for developers and operators," Merriman says. "[MySQL] doesn't have all those properties."

So they set out to build a database of their own, one that would discard the familiar relational database model in favor of a distributed platform tailored for moden-day web applications. "By reducing transactional semantics, we could still solve an interesting set of problems, but we could also scale," Merriman explains. After a year of work, the database was in place, and they decided it had as much potential as the cloud service it was designed for – if not more. The cloud service was never finished. But the database was open sourced as MongoDB.

Dwight Merriman and his team, including ShopWiki founder Eliot Horowitz, built MongoDB under the aegis of the New York–based startup 10gen, and the company now offers support, training, and consulting services for the database in addition to serving as the open source project's primary steward. This week, 10gen held its second annual San Francisco developer conference, and with his Tuesday-morning keynote, Merriman described the origins of MongoDB and explained why the database was built the way it was.

The split from the relational model was essential, Merriman says, because you can't do distributed joins in a way that readily scales. "I'm not smart enough to do distributed joins that scale horizontally, widely, and are super fast. You have to choose something else," Merriman explained during his talk. "We have no choice but to not be relational."

It was equally important, he says, to limit the database's transactional semantics. "You can do distributed transactions, but if you do them with no loss of generality and you do them across a thousand machines, it's not going to be that fast."

Intelligent flash storage arrays

Next page: Get off the Table

More from The Register

next story
Just don't blame Bono! Apple iTunes music sales PLUMMET
Cupertino revenue hit by cheapo downloads, says report
The DRUGSTORES DON'T WORK, CVS makes IT WORSE ... for Apple Pay
Goog Wallet apparently also spurned in NFC lockdown
IBM, backing away from hardware? NEVER!
Don't be so sure, so-surers
Hey - who wants 4.8 TERABYTES almost AS FAST AS MEMORY?
China's Memblaze says they've got it in PCIe. Yow
Microsoft brings the CLOUD that GOES ON FOREVER
Sky's the limit with unrestricted space in the cloud
This time it's SO REAL: Overcoming the open-source orgasm myth with TODO
If the web giants need it to work, hey, maybe it'll work
'ANYTHING BUT STABLE' Netflix suffers BIG Europe-wide outage
Friday night LIVE? Nope. The only thing streaming are tears down my face
Google roolz! Nest buys Revolv, KILLS new sales of home hub
Take my temperature, I'm feeling a little bit dizzy
Storage array giants can use Azure to evacuate their back ends
Site Recovery can help to move snapshots around
prev story

Whitepapers

Why and how to choose the right cloud vendor
The benefits of cloud-based storage in your processes. Eliminate onsite, disk-based backup and archiving in favor of cloud-based data protection.
Forging a new future with identity relationship management
Learn about ForgeRock's next generation IRM platform and how it is designed to empower CEOS's and enterprises to engage with consumers.
Reg Reader Research: SaaS based Email and Office Productivity Tools
Read this Reg reader report which provides advice and guidance for SMBs towards the use of SaaS based email and Office productivity tools.
Saudi Petroleum chooses Tegile storage solution
A storage solution that addresses company growth and performance for business-critical applications of caseware archive and search along with other key operational systems.
Protecting users from Firesheep and other Sidejacking attacks with SSL
Discussing the vulnerabilities inherent in Wi-Fi networks, and how using TLS/SSL for your entire site will assure security.