Feeds

Internet Archive stares down Google book mine

Beware the cookie before dinner

Top three mobile application threats

Google Books Settlement Con As part of his ongoing campaign against Google's $125m book-scanning pact, the Internet Archive's Peter Brantley has warned that even if authors opt-out of Google's Book Search service, the web giant will still have the power to mine their book data for use in other services.

"There is value in the comprehensiveness [of Google's digital collection of books]. It's something we need to think about, that scholars and researchers need to think about: whether or not they want to entrust this single comprehensive collection to a single corporate entity," Brantley said Friday during a conference dedicated to the Google Book Settlement at the University of California, Berkeley.

"Authors can remove works [from Google Book Search]. But Google will still be able to utilize those works, in many cases, for integration into other products, like Google Maps and others, that draw upon the variety of knowledge held in those books.

"As a rights holder, I can't pull my book from the corpus for data use only, and even if I pull my book for my display purposes, it's still in [the data corpus]."

In October, Google settled a lawsuit from the US Authors Guild and the Association of American Publishers over Book Search, a project that seeks to digitize the works inside the world's libraries. The pact creates a "Book Rights Registry" where authors and publishers can resolve copyright claims in exchange for a predefined cut of Google's revenues.

Brantley and so many others have complained that the settlement also gives Google a unique license to digitize and sell and post ads against "orphan works," books whose rights are controlled by authors and publishers who have yet to be located. But this is but one of Brantley's the concerns over the pact, which still requires court approval.

The settlement wouldn't just give Google a monopoly around the online display of orphan works. It would give the company the ability to freely mine the data inside the 10 million books (and counting) it has scanned - including books whose rights are held by authors and publishers who've opted out of Google Book Search.

Under the settlement, Google would offer others access to at least a portion of its collective book data - dubbed the "research corpus" - from inside two US-based libraries. Yes, only two. And the company could prevent others from using the data in what it views as competing services.

Dan Clancy, the engineering director of Google Book Search, said that Google would only crack down on services that compete with current Google offerings, not future offerings. But Google decides what competes and what doesn't.

Clancy warned, however, that if the settlement is rejected, only Google would have access to the corpus. This is the double edge of Google's Book Settlement. If it's approved, Google has an enormous amount of control over the future of digital books. And if it's rejected, the company still has an enormous amount of control. What's needed is federal legislation that addresses these (and so many other) issues. And that's what Brantley is pushing for.

"We have a responsibility to create a plethora of options. There will be a diversity of possible uses of this material that we can't possibly images," he said. "As scholars, do we think about alternatives. Do we think about other ways of working towards a future where this corpus might be available under better terms. We don't have to grab that cookie that's offered to us before dinner." ®

Build a business case: developing custom apps

More from The Register

next story
BBC goes offline in MASSIVE COCKUP: Stephen Fry partly muzzled
Auntie tight-lipped as major outage rolls on
iPad? More like iFAD: We reveal why Apple fell into IBM's arms
But never fear fanbois, you're still lapping up iPhones, Macs
Amazon Reveals One Weird Trick: A Loss On Almost $20bn In Sales
Investors really hate it: Share price plunge as growth SLOWS in key AWS division
Bose says today is F*** With Dre Day: Beats sued in patent battle
Music gear giant seeks some of that sweet, sweet Apple pie
There's NOTHING on TV in Europe – American video DOMINATES
Even France's mega subsidies don't stop US content onslaught
You! Pirate! Stop pirating, or we shall admonish you politely. Repeatedly, if necessary
And we shall go about telling people you smell. No, not really
Too many IT conferences to cover? MICROSOFT to the RESCUE!
Yet more word of cuts emerges from Redmond
Chips are down at Broadcom: Thousands of workers laid off
Cellphone baseband device biz shuttered
Twitch rich as Google flicks $1bn hitch switch, claims snitch
Gameplay streaming biz and search king refuse to deny fresh gobble rumors
prev story

Whitepapers

Designing a Defense for Mobile Applications
Learn about the various considerations for defending mobile applications - from the application architecture itself to the myriad testing technologies.
Implementing global e-invoicing with guaranteed legal certainty
Explaining the role local tax compliance plays in successful supply chain management and e-business and how leading global brands are addressing this.
Top 8 considerations to enable and simplify mobility
In this whitepaper learn how to successfully add mobile capabilities simply and cost effectively.
Seven Steps to Software Security
Seven practical steps you can begin to take today to secure your applications and prevent the damages a successful cyber-attack can cause.
Boost IT visibility and business value
How building a great service catalog relieves pressure points and demonstrates the value of IT service management.