Privacy pathology: It's time for the users to gather a little data – evidence

If Sherlock was alive today, he’d pack a Pi next to pistol and pipe

Opinion Almost exactly a month ago, we noted a splendid piece of academic research into Google's data-gathering and consent practises.

According to the research paper by Trinity College Dublin computer science professor Douglas Leith, the company had been gathering far too much user data from core messaging apps, and the forensic analysis of the network flow extracted a well-deserved mea culpa from Mountain View.

Now, it's Amazon's turn to find itself being examined for privacy breaches. Once again, the well-honed tools of the scientific method were used by researchers to unmask a spy in our lives, in this case the Alexa voice interaction system, Skills (what Alexa calls apps), and the bidding system that sells keyword-driven advertisement slots to advertisers.

Talk to Alexa about something, the academics found, and the auction price for related advertising opportunities goes up.

The complexity of this system is matched by the lack of transparency over how it works, much as the charming simplicity of using Alexa – you just gab at it – hides how it works. If you fancy annoying the device, try asking it what version of its software it's running, what diagnostics it has available, whether there's a debug mode, and so on. No matter what you may know about fault-finding or configuring computers or mobile devices, you won't get anywhere.

And so, the researchers had to design a complicated melange of hardware and software, including the obligatory Raspberry Pi, to load test data into the system and monitor what happened as a result. Read the paper for the details [PDF]. They're uncommonly good: as the researchers say, many reports of this kind lack enough technical detail to encourage others to replicate or build on the work.

Time for a little forensic computer science

By now, there's little argument that a new, legitimate and pressing field of forensic computer science is evolving, that of discovering and characterizing abuses of personal privacy and consent.

We know that the beasts who track us and feast on our data can only be considered as enemies; they promise symbiosis and law-abiding honesty, while blocking any attempts to verify this.

The regulators are overwhelmed and underfunded, the politicians are glacially slow and hard of hearing. The cloud is obdurately opaque; even proprietary software is amenable to decompilation and analysis if you have the code, but what goes on beyond the API is a true trade secret.

Science is our only hope.

Here, at last, is something the big tech firms can't hide. We know that they're collecting all that data for a purpose, and that for it to be worthwhile that purpose must be manifest.

With the Alexa finding, it was skewing the advertising market; the signal the researchers found which proved lack of compliance. This is the same realization that dogs all the intelligence agencies of the world: if you use the information you've found, sharp eyes will notice.

So, as with any novel process or phenomenon, the scientist selects the tests and watches the results. Whether the process wants to be understood or not isn't part of the equation. As chemists, physicists and biologists share techniques and data sets just as much as they do individual findings, a priority of this new science of data privacy must be to recognise itself as a field and start to curate its knowledge and aim for collective process.

The extremely smart but largely inaccessible tools and techniques demonstrated by both the Google and Amazon researchers of late should become as standard as lab equipment. It'll take a while, and funding for those who try to characterize the misdeeds of the very rich is often peculiarly hard to come by, no matter how pressing the need.

Pitching in

Even so, as in the early days of any new science, there is room for the amateur to do good work. Some experiments need nothing more than you, your devices, a pinch of methodology and a basic grasp of statistics. You don't even need a Raspberry Pi.

Pick five things at random that you'd never buy or find of the slightest interest, like wheelbarrows, golf clubs, collectable pig-themed ceramics, Windows 11. Discuss them in five different ways online – or even just in earshot of smart devices.

Before, during and after, meticulously note the number of ads or unrequested content you see, and the percentage, if any, of those which touch on the test data.

Do this with application over a decent period, much as a Victorian parson would map the ecology of his local meadow over the seasons, and the patterns will emerge. They can't help it. It's what they do.

The internet itself is an admirable substitute for the Victorian postal system when it comes to correspondence with like minds and the dissemination of transactions. Many secrets will be revealed.

There is no Nobel prize for any sort of computing, let alone the new science of data privacy. There is the reward of using the miscreants' strengths as vulnerabilities against themselves, of substituting the conscience that they lack with the knowledge that they're being watched, and of making the internet a less wretched, more honest place.

Fair exchange for spending five minutes talking about ceramic pigs. ®

Other stories you might like

  • Lonestar plans to put datacenters in the Moon's lava tubes
    How? Founder tells The Register 'Robots… lots of robots'

    Imagine a future where racks of computer servers hum quietly in darkness below the surface of the Moon.

    Here is where some of the most important data is stored, to be left untouched for as long as can be. The idea sounds like something from science-fiction, but one startup that recently emerged from stealth is trying to turn it into a reality. Lonestar Data Holdings has a unique mission unlike any other cloud provider: to build datacenters on the Moon backing up the world's data.

    "It's inconceivable to me that we are keeping our most precious assets, our knowledge and our data, on Earth, where we're setting off bombs and burning things," Christopher Stott, founder and CEO of Lonestar, told The Register. "We need to put our assets in place off our planet, where we can keep it safe."

    Continue reading
  • Conti: Russian-backed rulers of Costa Rican hacktocracy?
    Also, Chinese IT admin jailed for deleting database, and the NSA promises no more backdoors

    In brief The notorious Russian-aligned Conti ransomware gang has upped the ante in its attack against Costa Rica, threatening to overthrow the government if it doesn't pay a $20 million ransom. 

    Costa Rican president Rodrigo Chaves said that the country is effectively at war with the gang, who in April infiltrated the government's computer systems, gaining a foothold in 27 agencies at various government levels. The US State Department has offered a $15 million reward leading to the capture of Conti's leaders, who it said have made more than $150 million from 1,000+ victims.

    Conti claimed this week that it has insiders in the Costa Rican government, the AP reported, warning that "We are determined to overthrow the government by means of a cyber attack, we have already shown you all the strength and power, you have introduced an emergency." 

    Continue reading
  • China-linked Twisted Panda caught spying on Russian defense R&D
    Because Beijing isn't above covert ops to accomplish its five-year goals

    Chinese cyberspies targeted two Russian defense institutes and possibly another research facility in Belarus, according to Check Point Research.

    The new campaign, dubbed Twisted Panda, is part of a larger, state-sponsored espionage operation that has been ongoing for several months, if not nearly a year, according to the security shop.

    In a technical analysis, the researchers detail the various malicious stages and payloads of the campaign that used sanctions-related phishing emails to attack Russian entities, which are part of the state-owned defense conglomerate Rostec Corporation.

    Continue reading
  • FTC signals crackdown on ed-tech harvesting kid's data
    Trade watchdog, and President, reminds that COPPA can ban ya

    The US Federal Trade Commission on Thursday said it intends to take action against educational technology companies that unlawfully collect data from children using online educational services.

    In a policy statement, the agency said, "Children should not have to needlessly hand over their data and forfeit their privacy in order to do their schoolwork or participate in remote learning, especially given the wide and increasing adoption of ed tech tools."

    The agency says it will scrutinize educational service providers to ensure that they are meeting their legal obligations under COPPA, the Children's Online Privacy Protection Act.

    Continue reading
  • Mysterious firm seeks to buy majority stake in Arm China
    Chinese joint venture's ousted CEO tries to hang on - who will get control?

    The saga surrounding Arm's joint venture in China just took another intriguing turn: a mysterious firm named Lotcap Group claims it has signed a letter of intent to buy a 51 percent stake in Arm China from existing investors in the country.

    In a Chinese-language press release posted Wednesday, Lotcap said it has formed a subsidiary, Lotcap Fund, to buy a majority stake in the joint venture. However, reporting by one newspaper suggested that the investment firm still needs the approval of one significant investor to gain 51 percent control of Arm China.

    The development comes a couple of weeks after Arm China said that its former CEO, Allen Wu, was refusing once again to step down from his position, despite the company's board voting in late April to replace Wu with two co-chief executives. SoftBank Group, which owns 49 percent of the Chinese venture, has been trying to unentangle Arm China from Wu as the Japanese tech investment giant plans for an initial public offering of the British parent company.

    Continue reading
  • SmartNICs power the cloud, are enterprise datacenters next?
    High pricing, lack of software make smartNICs a tough sell, despite offload potential

    SmartNICs have the potential to accelerate enterprise workloads, but don't expect to see them bring hyperscale-class efficiency to most datacenters anytime soon, ZK Research's Zeus Kerravala told The Register.

    SmartNICs are widely deployed in cloud and hyperscale datacenters as a means to offload input/output (I/O) intensive network, security, and storage operations from the CPU, freeing it up to run revenue generating tenant workloads. Some more advanced chips even offload the hypervisor to further separate the infrastructure management layer from the rest of the server.

    Despite relative success in the cloud and a flurry of innovation from the still-limited vendor SmartNIC ecosystem, including Mellanox (Nvidia), Intel, Marvell, and Xilinx (AMD), Kerravala argues that the use cases for enterprise datacenters are unlikely to resemble those of the major hyperscalers, at least in the near term.

    Continue reading

Biting the hand that feeds IT © 1998–2022