Who's archiving IT's history?

The mishandling of Babbage's baggage


Column The relaunch of the IT History Society (formerly the Charles Babbage Foundation) exposes a problem which the Web has brought to journalism and historians - stuff is not being preserved. There are people, however, who are trying to build proper records of the past.

The question is whether this works. The IT History Society is trying to create a history of European computing:

Software for Europe is a historical project with strong interdisciplinary connections. Its members are informed variously by the disciplinary perspectives of cultural history, business history, economic history, history of science and technology, science and technology studies, and technology policy; it is our intention that the project’s work will be recognised in all these fields. Much of the work will proceed from analysis of written sources, published and archival, and to some extent on the examination of software itself.

Well, they're in for a tough job. One of my jobs recently has been to look back into IT history and apply some 20-20 hindsight to events five years ago and ten years ago.

How hard can that be? I go into the office; I open the vault with back numbers of IT Week. I find the one with the same date as my next publication deadline, and I flip through the pages till I find something interesting and topical. And I write my column. Bingo.

Last week, I couldn't get into the office, so I went for my backup. Way back up: the WayBackMachine which is attempting to build up an archive of the web.

I won't hear a word against the WayBackMachine. But I will in honesty have to say a few words against it: it's got holes.

What it's good at is holding copies of "That day's edition" just the way a newspaper archive does. I can, for example, go back to NewsWireless by opening up this link; and there, I can find everything that was published on December 6th 2002 - five years ago! - more or less. I can even see that the layout was different, if I look at the story of how NewsWireless installed a rogue wireless access point in the Grand Hotel Palazzo Della Fonte in Fiuggi, in the hills above Rome. And it shows that in those early days, NewsWireless was called "Guy Kewney's Mobile Campaign" - a name we chose in spite of the good advice of more sensible friends...

Now, have a look at the same story, as it appears on NewsWireless today. The words are there, but it looks nothing like it used to look.

Unusually, NewsWireless does give you the same page you would have seen five years ago. When you're reading the Fiuggi story, the page shows you contemporary news: for example the utterly mad, definitely disgusting "phones to salivate over" http://www.newswireless.net/index.cfm/article/1226 which appeared that week. It's the week's edition, in content at least.

Most websites don't do this.

You can, sometimes, track back a particular five-year-old story (though sadly you'll often find it's been deleted), but if you go to the original site you're likely to find that the page you see is surrounded by modern stories. It's not a five-year-old edition. Take, for example Gordon Laing's Christmas 2002 article on five megapixel super-cameras (I know, I know) and you'll find exactly no stories at all relating to Christmas 2002. They were published, yes, but they aren't archived together anywhere - except the WayBackMachine.

And sadly, the WayBackMachine has a limitation; it has a lot of historical "editions" of websites, but there are a lot more it doesn't have.

Also, it's vulnerable to electronic interference. Censorshipware, according to Seth Finkelstein, sees the archive as a single site. If just one page appears to contain an image showing more skin than a maiden aunt in Boise, Idaho, might consider proper, then the entire archive is liable to be blacked by a lot of ISPs who take their cue from that net nanny.


Other stories you might like

  • HPE unveils Arm-based ProLiant server for cloud-native workloads
    Looks like it went with Ampere – which means a certain Reg writer lost a bet

    Arm has a champion in the shape of HPE, which has added a server powered by the British chip designer's CPU cores to its ProLiant portfolio, aimed at cloud-native workloads for service providers and enterprise customers alike.

    Announced at the IT titan's Discover 2022 conference in Las Vegas, the HPE ProLiant RL300 Gen11 server is the first in a series of such systems powered by Ampere's Altra and Altra Max processors, which feature up to 80 and 128 Arm-designed Neoverse cores, respectively.

    The system is set to be available during Q3 2022, so sometime in the next three months, and is basically an enterprise-grade ProLiant server – but with an Arm CPU at its core instead of the more usual Intel Xeon or AMD Epyc X86 chips.

    Continue reading
  • US weather forecasters power up latest supercomputers to keep you out of the rain
    NOAA makes it rain for HPE, AMD

    Predicting the weather is a notoriously tricky enterprise, but that’s never held back America's National Oceanic and Atmospheric Administration (NOAA). After more than two years of development, the agency brought a pair of supercomputers online this week that it says will enable more accurate forecast models.

    Developed and maintained by General Dynamics Information Technology (GDIT) under an eight-year contract, the Cactus and Dogwood supers — named after the fauna native to the machines' homes in Phoenix, Arizona, and Manassas, Virginia, respectively — will support larger, higher-resolution models than previously possible. The cost to build, house, and support and operate these machines, now operational, will cost $150 million over the next five years, we understand.

    “People are looking for the best possible weather forecast information that they can get,” Brian Gross, director of the Environmental Modeling Center for the National Weather Service, told The Register.

    Continue reading
  • Google said to be taking steps to keep political campaign emails out of Gmail spam bin
    Just after Big Tech comes under fire for left and right-leaning message filters

    Google has reportedly asked the US Federal Election Commission for its blessing to exempt political campaign solicitations from spam filtering.

    The elections watchdog declined to confirm receiving the supposed Google filing, obtained by Axios, though a spokesperson said the FEC can be expected to publish an advisory opinion upon review if Google made such a submission.

    Google did not immediately respond to a request for comment. If the web giant's alleged plan gets approved, political campaign emails that aren't deemed malicious or illegal will arrive in Gmail users' inboxes with a notice asking recipients to approve continued delivery.

    Continue reading
  • China is trolling rare-earth miners online and the Pentagon isn't happy
    Beijing-linked Dragonbridge flames biz building Texas plant for Uncle Sam

    The US Department of Defense said it's investigating Chinese disinformation campaigns against rare earth mining and processing companies — including one targeting Lynas Rare Earths, which has a $30 million contract with the Pentagon to build a plant in Texas.

    Earlier today, Mandiant published research that analyzed a Beijing-linked influence operation, dubbed Dragonbridge, that used thousands of fake accounts across dozens of social media platforms, including Facebook, TikTok and Twitter, to spread misinformation about rare earth companies seeking to expand production in the US to the detriment of China, which wants to maintain its global dominance in that industry. 

    "The Department of Defense is aware of the recent disinformation campaign, first reported by Mandiant, against Lynas Rare Earth Ltd., a rare earth element firm seeking to establish production capacity in the United States and partner nations, as well as other rare earth mining companies," according to a statement by Uncle Sam. "The department has engaged the relevant interagency stakeholders and partner nations to assist in reviewing the matter.

    Continue reading
  • California's attempt to protect kids online could end adults' internet anonymity
    Websites may be forced to verify ages of visitors unless changes made

    California lawmakers met in Sacramento today to discuss, among other things, proposed legislation to protect children online. The bill, AB2273, known as The California Age-Appropriate Design Code Act, would require websites to verify the ages of visitors.

    Critics of the legislation contend this requirement threatens the privacy of adults and the ability to use the internet anonymously, in California and likely elsewhere, because of the role the Golden State's tech companies play on the internet.

    "First, the bill pretextually claims to protect children, but it will change the Internet for everyone," said Eric Goldman, Santa Clara University School of Law professor, in a blog post. "In order to determine who is a child, websites and apps will have to authenticate the age of ALL consumers before they can use the service. No one wants this."

    Continue reading
  • Is computer vision the cure for school shootings? Likely not
    Gun-detecting AI outfits want to help while root causes need tackling

    Comment More than 250 mass shootings have occurred in the US so far this year, and AI advocates think they have the solution. Not gun control, but better tech, unsurprisingly.

    Machine-learning biz Kogniz announced on Tuesday it was adding a ready-to-deploy gun detection model to its computer-vision platform. The system, we're told, can detect guns seen by security cameras and send notifications to those at risk, notifying police, locking down buildings, and performing other security tasks. 

    In addition to spotting firearms, Kogniz uses its other computer-vision modules to notice unusual behavior, such as children sprinting down hallways or someone climbing in through a window, which could indicate an active shooter.

    Continue reading

Biting the hand that feeds IT © 1998–2022