Feeds

British Library wants taxpayer to gobble the web

Cost? We don't know

Build a business case: developing custom apps

British Library wants to archive the UK web, creating an invaluable national treasure trove of porn, celebrity trivia gossip and Daily Mail comments. But it admits it can't put a figure on the project - which looks like becoming a huge, open-ended commitment for the taxpayer.

Today the Library stepped up the pressure for the law to be changed, allowing copyright libraries to create copies of web material for research purposes of other copyright holders material. Five statutory libraries already have permission to make printed material available. Now the British Library says it wants the Web too.

"It's not a request for additional funding," a BL spokesperson said, but they couldn't say how much the creeping mission would end up costing us. At first, the BL won't archive every Tweet, but do an annual crawl, with some sites such as No 10 Downing Street archived more often. That would cost 220TB of data, it reckons about £4,000 in storage.

But that would barely make a dimple in a replica of UK web output, now that so many non-web chat areas have migrated to a home between angle brackets. The BL acknowledges there are eight million sites.

What, we wondered, was the point of archiving every single "Ashlee Cole iz a slag" typed into a browser?

"It may be that somebody wants to look back and research celebrity and this could be important to their research," we were told.

No doubt. But every Tweet and comment?

It was cheaper, the spokesman assured us, than employing a curator to choose between the best Ashley/Cheryl comments (for example).

Ah, right. So the mechanics dictate the curation policy.

But it was also fairer, he added, because the neutral, objective web bot couldn't be accused of bias. Even in momentous national conversations as the Cole divorce.

There are plenty of comments flying around this morning wondering why public money should be required to archive more than a handful of websites. Especially with Brewster Kahle's Archive.Org, which is privately funded.

At first the library told us the public was unaware that websites disappear without some part of the British state keeping a copy - an interesting claim. I've never met anyone who thinks all websites are preserved by some silent, omniscient backup programme.

Then the Library told us that the private sector couldn't be trusted to do the job, because future funding couldn't be assured. But with the British state in the red to the tune of £180bn this year, a defecit larger than Greece's in GDP terms (12.8 per cent), and frontline services such as nurses facing the chop, it's questionable whether anyone wants prefers to keep a copy of those Mail comments instead. ®

Build a business case: developing custom apps

More from The Register

next story
Assange™: Hey world, I'M STILL HERE, ignore that Snowden guy
Press conference: ME ME ME ME ME ME ME (cont'd pg 94)
Premier League wants to PURGE ALL FOOTIE GIFs from social media
Not paying Murdoch? You're gonna get a right LEGALLING - thanks to automated software
Online tat bazaar eBay coughs to YET ANOTHER outage
Web-based flea market struck dumb by size and scale of fail
Amazon takes swipe at PayPal, Square with card reader for mobes
Etailer plans to undercut rivals with low transaction fee offer
US regulators OK sale of IBM's x86 server biz to Lenovo
Now all that remains is for gov't offices to ban the boxes
XBOX One will learn to play media from USB and DLNA sources
Hang on? Aren't those file formats you hardly ever see outside torrents?
Class war! Wikipedia's workers revolt again
Bourgeois paper-shufflers have 'suspended democracy', sniff unpaid proles
prev story

Whitepapers

Endpoint data privacy in the cloud is easier than you think
Innovations in encryption and storage resolve issues of data privacy and key requirements for companies to look for in a solution.
Implementing global e-invoicing with guaranteed legal certainty
Explaining the role local tax compliance plays in successful supply chain management and e-business and how leading global brands are addressing this.
Top 8 considerations to enable and simplify mobility
In this whitepaper learn how to successfully add mobile capabilities simply and cost effectively.
Solving today's distributed Big Data backup challenges
Enable IT efficiency and allow a firm to access and reuse corporate information for competitive advantage, ultimately changing business outcomes.
Reg Reader Research: SaaS based Email and Office Productivity Tools
Read this Reg reader report which provides advice and guidance for SMBs towards the use of SaaS based email and Office productivity tools.