Feeds

Blog noise achieves Google KO

Can we rescue our search engine from the wiki-fiddlers?

  • alert
  • submit to reddit

Security for virtualized datacentres

"All weblogs have their useful side
it's the men who write the weblog software I can't abide." - Bertolt Brecht [*]

The humble weblog has finally achieved dominance over Google, the world's most-used search engine. Originally intended as a tool that allowed people to publish their personal diaries, weblog software has swiftly evolved, accreting several "innovations" that have had catastrophic consequences for Google. If you've never heard of the "Trackback", or ever wanted to know, then we have bad news: you're about to become acquainted, whether you like it or not, dear Google user.

A "Trackback" is an auto-citation feature that allows solitary webloggers to feel as if they are part of a community. It's a cunning trick that allows the reader to indicate that they've read a weblog entry, or as the official description from MovableType has it: "Using TrackBack, the other weblogger can automatically send a ping to your weblog, indicating that he has written an entry referencing your original post."

The original blog then sprouts a list of "trackback" entries from other webloggers who have read, and linked to the original article. Kinda neat, huh? Except for one unforeseen technical consequence: the Trackback generates an empty page, and Google - being too dumb to tell an empty page from the context that surrounds it - gives it a very high value when it calculates its search results. So Google's search results are littered with empty pages.

Try this for size: it's a Google query for OS X Panther discussion. In what must be a record, Google is - at time of writing - returning empty Trackback pages as No.1, No.2, No.3 and No.4 positions. No.5 gets you to a real web page - an Apple Insider bulletin board. Then it's back to empty Trackback pages for results No.6, No.7 and No.10. In short, Google returns blog-infested blanks for seven of the top entries.

So who's to blame?

Well, let's assume that Google in good faith simply wants to give us good search results. Not because it's more good, or any less corrupt than any other secretive California corporation, but because it knows that useless searches - as the one we have just described - will repel users. As one reader pointed out: blog noise means life or death for Google.

So suspicions fall on the weblog tools vendors, who have unleashed such a potent toxin that it renders the world's leading search engine a dud. Who are they, exactly?

One reader uncharitably points to a class of 'wiki-wankers' - a term possibly too rude by far for us - meaning the now-unemployed generation of dotcom-era HTML coders who have undoubted skill at producing whizzy, haiku-length hacks, but who can't see the ecological consequences of their own actions. We're pretty sure that "Fill Google with empty pages" wasn't on the Trott's to-do list.

But off they went, and here we are.

When the old longhair database people, now barely remembered, went off and designed information systems, they thought of values such as data integrity and resilience. This created a rigorous and unforgiving peer-review culture, but the values survived. Lacking such peer review, today's wiki-fiddlers can create such catastrophes as Trackbacks with apparent impunity.

In fact publisher Tim O'Reilly cited the Trackback as the greatest innovation of the "Emerging Technology" conference of 2002, before going on to advocate the use of Cascading Style Sheets as a transmission protocol in the aftermath of this year's conference. In such circumstances, you have to conclude that no-one is minding the wiki-fiddlers' playpen.

In the face of concerted attacks to undermine its integrity from link farms and webloggers Google has taken drastic remedial action: abandoning PageRank™ and instating some brutal emergency filters. Time will tell if it can succeed. ®

[*] Very nearly: "All vices have their useful side - it's the men who practice them I can't abide" - Baal

Choosing a cloud hosting partner with confidence

More from The Register

next story
The 'fun-nification' of computer education – good idea?
Compulsory code schools, luvvies love it, but what about Maths and Physics?
Ex-US Navy fighter pilot MIT prof: Drones beat humans - I should know
'Missy' Cummings on UAVs, smartcars and dying from boredom
Facebook, Apple: LADIES! Why not FREEZE your EGGS? It's on the company!
No biological clockwatching when you work in Silicon Valley
Happiness economics is bollocks. Oh, UK.gov just adopted it? Er ...
Opportunity doesn't knock; it costs us instead
'Cowardly, venomous trolls' threatened with TWO-YEAR sentences for menacing posts
UK government: 'Taking a stand against a baying cyber-mob'
Sysadmin with EBOLA? Gartner's issued advice to debug your biz
Start hoarding cleaning supplies, analyst firm says, and assume your team will scatter
Doctor Who's Flatline: Cool monsters, yes, but utterly limp subplots
We know what the Doctor does, stop going on about it already
prev story

Whitepapers

Forging a new future with identity relationship management
Learn about ForgeRock's next generation IRM platform and how it is designed to empower CEOS's and enterprises to engage with consumers.
Cloud and hybrid-cloud data protection for VMware
Learn how quick and easy it is to configure backups and perform restores for VMware environments.
Three 1TB solid state scorchers up for grabs
Big SSDs can be expensive but think big and think free because you could be the lucky winner of one of three 1TB Samsung SSD 840 EVO drives that we’re giving away worth over £300 apiece.
Reg Reader Research: SaaS based Email and Office Productivity Tools
Read this Reg reader report which provides advice and guidance for SMBs towards the use of SaaS based email and Office productivity tools.
Security for virtualized datacentres
Legacy security solutions are inefficient due to the architectural differences between physical and virtual environments.