104 comments

colonwqbang1 day ago
When you think about it, 24 years is not such a long time for a machine to keep doing what it's supposed to be doing. It says something about our profession that we find a 24-year service life surprising.

There are airliners that have been in service for over 50 years.

VectorLock1 day ago
>There are airliners that have been in service for over 50 years.

Are there any airliners that have flown for 24 years non-stop?

serf1 day ago
this is why I find comparing mostly-solid-state stuff with big mechanical contraptions a frustrating exercise.

it's more impressive to me that the electric fans within the Stratus server still operate than it is that the chips still work, and if we're going with the plane metaphor it's the silicon that puts the thing into the air, not the simple fans.

So really it just boils down to "well, the planes work is harder." as to why it fails, which is of course abstract and unsatisfying; comparing the comparative lifetime workloads of a chip that has shifted trillions of bits versus some jet engines that have moved millions of kilograms around the world.

and then another layer comes and makes it even weirder : the vast majority of hardware around the world that has been EoLd and replaced has been in that situation due to software, not the hardware itself; a paradigm that really doesn't exist in the physical engineering world in the same was as it does CS.

rurbanabout 19 hours ago
You'll find for sure an old DC3 somewhere, and a Cessna 172 also. And according to https://www.oldest.org/technology/planes-still-flying/ there are still 40 year old 737's in the air
dlisboa1 day ago
Yeah, it’s not a good comparison. Give me endless parts and a maintenance window of 6 months every so often and I can make any server last a century.
That would not be a useful airliner.
Was going to say, there's servers that last several decades, they receive routine maintenance, most people buy a computer, and rarely replace parts, and if enough years pass, you opt to buy a new one instead.
tyrabound1 day ago
You raised in my mind that we are looking at this all wrong, the server is not actually one component, it’s really a system of systems, and just like an airliner is a system of systems too, the conflict arises in that the comparison is at the wrong level.

It seems when people think of a server today, they think of a single computer, if not some software server somewhere in the cloud. These subject kinds of servers are far closer to complex systems of separate components, more like a network system, it’s why components of what are really separate networked computers can and need to be replaced.

An airliner is of course also made up of components that are also systems, however at least in my mind, the difference is the actual expected use case. An airliner is not a good comparison because it is never expected to remain in continuous operation, e.g., that at least one engine is always running even when it is being overhauled or repaired, to satisfy a requirement of continuous operation.

pitched1 day ago
FTA, this is a fault-tolerant server where all hardware is redundant and can be hot swapped while running. The servers you’re thinking about are redundant at the software level so rebooting one server won’t cause the service to go down.

What is remarkable is that apparently that VOS thing has never crashed. I wonder if they also do software redundancy under the hood to keep uptime going during reboots. If the CPU is hotswappable, it must have something.

afavour1 day ago
I think it just speaks to how quickly the tech has developed. Airliners work more or less the same as they did 50 years ago, computers are vastly different (and vastly more powerful).
Leonard_of_Q1 day ago
Airliners may look the same but they don't work as they did 50 years ago when the was no fly by wire in these planes, no high-bypass engines, etc.
asah1 day ago
yabut... airliners require constant maintenance and periodic rebuilds.
Not to mention service disruptions and unplanned downtime.
bitwize1 day ago
One of the major differences between like, your PC, and a high-availability server is that high-availability servers require those things as well—and get them, piecewise, while still running. It's like doing a complete overhaul while the plane is flying.

IBM mainframes, famously, phone home if they detect a faulty component. Service personnel will be on site, same day, to replace it before you even knew it was there, let alone had the opportunity to ask.

lp921 day ago
Those aircraft have maintainance windows/refurbishment/refits and aren't flying 24/7 though.
debo_1 day ago
This server also had planned downtime maintenance. The article specifically mentions "no unplanned downtime."
mech4221 day ago
I started working on Stratus machines right out of high school... They took me from backwoods New England to Wall St.

The were Awesome machines - don't let those filthy hobbitsis...err..tandem users cloud your judgement. Stratus was pure hardware FT, and tandem... wasn't :-P

Vos was pretty nice too - decent macro language - not all the utils you expect with a unix box, but in terms of the core OS and shell it was pretty good machine .. 40 years later I still love those machines and the career they helped get me into :-D

mkovach1 day ago
I spent six months programming on VOS after a gig on a VAX. Yes, I'll be outside telling the kids to get off my lawn as soon as I find my reading glasses and have a nap.

VOS was intuitive and did what I needed. Nothing flashy. Just useful. I miss that sometimes. Give me a tool that works; if it also happens to be great, lovely.

I didn't work much with the Stratus hardware, but reading about it was fascinating. I wish I remembered more of it.

keiferwiseman1 day ago
Oh Tandem, you don’t hear that much anymore. They still have plenty of them still running banks and what not under the name HP non-stop.
mech4221 day ago
yeah - the big difference was Tandem was software (+ hardware) fault tolerant while Stratus was full hardware level fault tolerant. This gave Tandem a real price edge.

Both were really popular with banks and brokerage houses (back when I worked with stratus at brokerages, HFT was just emerging and down time was umm... frowned upon :-D

arscan1 day ago
Hey, I worked there in high school as an intern! But just on the intranet, which for its time (‘96?) was really quite good.
nineteen9991 day ago
We use to run Stratus VOS on Continuum series hardware at a lottery/poker machine company I worked at in the mid 2000's. It handled the jackpot system which decided which of 13,500 poker machines in our state would jackpot at any given time.

A reliable beast - never failed once in the entire time I worked there. Based on PA-RISC if I recall correctly.

jll291 day ago
Great story, thanks. I wish it said more about the architecture and also VOS.

Speaking of reliable beasts, I own(ed) a HP PA-RISC workstation that never crashed in six years, and it was even purchased refurbished (second hand, and educational rebate, and still the cost of a car in 1996 - but worth every penny).

I'd add "R.I.P. HP 9000 715/75" but of course she is STILL working now, twenty years later, and she is sitting right next to me (also got a second 715/100XC for spare parts in case on day they are needed).

sdcfgy1 day ago
The old workstations were always pretty damn reliable. I’ve got a Sun Ultra 30 in storage. I may grab it and boot it up and see if it still works. Suspect the NVRAM and power supply caps are duff though.
iberator1 day ago
what are you using it with? hp unix? netbsd?
mech4221 day ago
IIRC, Stratus was M68K (like 12 per cpu board organized into 3 logical cpu's for fault checking)
kjs31 day ago
The OG Stratus FT was mc68k; what a beast. They moved to Intel i860 in the early 90s (XA series...I had a processor board from one hanging in my conference room back in the day), then HP-PA (Continuum series) in the 2000s and these days it's apparently virtualized on x86. Still very popular in the credit card processing world.
zorked1 day ago
The poker machines are connected? Why wouldn't they just jackpot locally and indepdendently?
rafaelvasco1 day ago
Can't have several machines "jackpotting" at the same time or similar time. The prizing algorithm is very strict.
toast01 day ago
Networked jackpots are a user visible and aparently user desirable feature.

When the jackpot is fed by all the machine in the state, they can grow much faster.

voidUpdate1 day ago
> "This system runs an older version Stratus proprietary VOS operating system, which Hogan believes hasn’t been updated since the early 2000s. “It’s been extremely stable,’ he said."

That's not particularly surprising to be honest. If it's a server operating system, on hardware that doesn't change, there's not really a reason to update the OS unless there's a security problem found

mrweasel1 day ago
It's not really a security strategy, but running an obscure proprietary operating system on a system which probably isn't connected to the internet means that you're only susceptible to highly targeted attacks by a dedicated foe.

Apparently the latest VOS release was in 2025, which means that somewhere there's a team of developers maintaining it. That's really the thing I find the most interesting, that there exists teams developing operating systems like VOS, strange DOS clones or UnixWare (perhaps less so these days). Must be a strange anonymous life.

> Must be a strange anonymous life.

This perfectly characterizes most embedded systems work. Competent software development using bespoke or off the shelf components that are not mainstream. Nobody ever realizes they’re using your computer software if you’ve done your job right. One day someone throws your appliance away because it’s obsolete.

voidUpdate1 day ago
The article is very vague about what it actually does, so I don't know if it's connected to the internet or not :/
nathell1 day ago
Reminds me of a C64C that, as of 2016, was used in production in a car workshop in Poland:

https://www.neoteo.com/en/a-commodore-64-stands-the-test-of-...

Read the full thread on Hacker News →

Related stories