Feeds

Lateral thought saves sizzling server

Game, set and crash

High performance access to file storage

D'oh! I learned a long time ago that generating random numbers (really, truly random numbers) is a non-trivial exercise.

However, I completely failed to apply that computer science lesson to the real world of computing and continued to believe that events in the Newtonian world could happen without a cause. Such a belief system is not usually dangerous but, when applied to solving computer problems, it can be a serious disadvantage.

I came to my senses after I had spent a great deal of time trying to track down an intermittent fault on a NetWare server. It crashed. Then it ran fine for days. And then it crashed again. And again. Apparently, at random.

We would get a call from the local supervisor Rosanne whenever it crashed. By the time we got there (it was offsite) the server would reboot as if nothing had happened and run like a dream.

Sometimes it would run for weeks, other times it crashed three days running. The only correlation we could spot was that the crashes were always during the day so in that sense it wasn't random but since days happen seven times a week, every week without fail, it wasn't really a great help in diagnosing the problem and curing it.

And during the day there was absolutely no correlation with load. We concluded that the server was crashing at random and started the process of swapping parts (at random!) to try to cure it.

Then, on one of our frequent visits, Rosanne said jokingly that we really had to fix the problem because it was ruining her social life. The server crashed every time she played tennis with her new boyfriend. By this time we were desperate to find any correlation between the crashes and real life so we rather startled her by resurrecting the Spanish inquisition.

Was she serious? Well, er... not every time but, yeah, her boyfriend had pulled her leg that she was setting off her pager on purpose to avoid losing games and she had realized that it did seem to happen all too frequently. How often did she play? Well, a couple of time a week, maybe; it depended.

On what? How did she decide to play? Well, both she and her boyfriend worked flexitime so whenever the weather was good, they booked a court and played a game. They made up the missing time by working an hour later in the evening.

Ace in the hole

So the server was crashing when the weather was good. OK, how do we define good weather in Scotland where this was all happening? It is good weather if the sun shines. What happens when the sun shines? The sky is bluer, there are fewer clouds, the humidity is probably lower... it gets hotter. Hmm.

Servers don't like heat. Where is the server? Sitting on a bench. In front of a south facing window - this was in the days before server rooms, when air conditioning was provided only for mainframes.

So, Rosanne plays tennis when the weather is good, the sun shines and it's cooking the server. The Newtonian world is back in balance, yin has a yang and effect does have a cause.

I am, I like to think, at least slightly wiser now. I learned from that particular lesson that saying: "It must be random" is another way of saying that I have yet to find the correlation. Worse than that, it's usually a cop out.®

Doh! Is Mark Whitehorn's look at the events, and lessons learned, that served him well during his computing career.

High performance access to file storage

More from The Register

next story
European Court of Justice rips up Data Retention Directive
Rules 'interfering' measure to be 'invalid'
Dropbox defends fantastically badly timed Condoleezza Rice appointment
'Nothing is going to change with Dr. Rice's appointment,' file sharer promises
Cisco reps flog Whiptail's Invicta arrays against EMC and Pure
Storage reseller report reveals who's selling what
This time it's 'Personal': new Office 365 sub covers just two devices
Redmond also brings Office into Google's back yard
Bored with trading oil and gold? Why not flog some CLOUD servers?
Chicago Mercantile Exchange plans cloud spot exchange
Just what could be inside Dropbox's new 'Home For Life'?
Biz apps, messaging, photos, email, more storage – sorry, did you think there would be cake?
IT bods: How long does it take YOU to train up on new tech?
I'll leave my arrays to do the hard work, if you don't mind
prev story

Whitepapers

Securing web applications made simple and scalable
In this whitepaper learn how automated security testing can provide a simple and scalable way to protect your web applications.
Five 3D headsets to be won!
We were so impressed by the Durovis Dive headset we’ve asked the company to give some away to Reg readers.
HP ArcSight ESM solution helps Finansbank
Based on their experience using HP ArcSight Enterprise Security Manager for IT security operations, Finansbank moved to HP ArcSight ESM for fraud management.
The benefits of software based PBX
Why you should break free from your proprietary PBX and how to leverage your existing server hardware.
Mobile application security study
Download this report to see the alarming realities regarding the sheer number of applications vulnerable to attack, as well as the most common and easily addressable vulnerability errors.