Redmond researchers profile Skype scammers

Markov models sniff out the fake IDs and spammy creeps


A group of Microsoft researchers has used supervised machine learning to try and improve detection of fraudulent user accounts.

With Skype as their test platform, the group says it was able to achieve 68 per cent successful detection of fake accounts within four months of activity, while keeping false positives down to 5 per cent. At the same time, it claimed, it achieved a reduction in “the number of undetected fraudulent users active for over 10 months by a factor of 2.3.”

The aim, the researchers say, was to identify fraudsters “that have eluded the first line of detection systems and have been active for months”, drawing on static profile information such as age, active profile information such as the time series of a user's calls, social behaviour (adding or deleting friend contacts), and social features (such as PageRank).

The researchers note that one of their statistical techniques, the application of hidden Markov models (HMMs) to this context is new (although, interestingly, one of the authors of the paper, Moises Goldszmidt, has previously used HMMs to help predict disk failure in data centres).

The research was based on a pool of 100,000 each of legitimate and fraudulent users (as nominated by Skype), which yielded a test pool of 34,000 accounts, based on accounts which existed for 4 months without being blocked.

To protect information about users whose account information was provided by Skype for the study, the paper states that:

  • All Skype IDs were anonymized using a one-way cryptographic salted hash function
  • The only usage data applied to the study was the number of days each month that a user accessed particular features, such as chat, Skype calls, video calls, and Skype In / Skype Out.
  • Data was kept on a dedicated machine with restricted access, and the researchers planned to delete the data when their research was complete.

Although this work concentrated on Skype, “we chose not to rely on Skype's informal intent in those definitions, nor on Skype's software … in order to develop robust, self-contained methods.”

In other words, if the work proved valuable deployed to Skype, the researchers hope the techniques could be applied to other platforms as well.

Lead author Anna Leontjeva, who was an intern at Microsoft Research when she conducted the research, is an Estonian student from the University of Tartu. The paper was prepared for last November's AISec’13 in Berlin, and is available here. ®

Broader topics


Other stories you might like

  • AI tool finds hundreds of genes related to human motor neuron disease

    Breakthrough could lead to development of drugs to target illness

    A machine-learning algorithm has helped scientists find 690 human genes associated with a higher risk of developing motor neuron disease, according to research published in Cell this week.

    Neuronal cells in the central nervous system and brain break down and die in people with motor neuron disease, like amyotrophic lateral sclerosis (ALS) more commonly known as Lou Gehrig's disease, named after the baseball player who developed it. They lose control over their bodies, and as the disease progresses patients become completely paralyzed. There is currently no verified cure for ALS.

    Motor neuron disease typically affects people in old age and its causes are unknown. Johnathan Cooper-Knock, a clinical lecturer at the University of Sheffield in England and leader of Project MinE, an ambitious effort to perform whole genome sequencing of ALS, believes that understanding how genes affect cellular function could help scientists develop new drugs to treat the disease.

    Continue reading
  • Need to prioritize security bug patches? Don't forget to scan Twitter as well as use CVSS scores

    Exploit, vulnerability discussion online can offer useful signals

    Organizations looking to minimize exposure to exploitable software should scan Twitter for mentions of security bugs as well as use the Common Vulnerability Scoring System or CVSS, Kenna Security argues.

    Better still is prioritizing the repair of vulnerabilities for which exploit code is available, if that information is known.

    CVSS is a framework for rating the severity of software vulnerabilities (identified using CVE, or Common Vulnerability Enumeration, numbers), on a scale from 1 (least severe) to 10 (most severe). It's overseen by First.org, a US-based, non-profit computer security organization.

    Continue reading
  • Sniff those Ukrainian emails a little more carefully, advises Uncle Sam in wake of Belarusian digital vandalism

    NotPetya started over there, don't forget

    US companies should be on the lookout for security nasties from Ukrainian partners following the digital graffiti and malware attack launched against Ukraine by Belarus, the CISA has warned.

    In a statement issued on Tuesday, the Cybersecurity and Infrastructure Security Agency said it "strongly urges leaders and network defenders to be on alert for malicious cyber activity," having issued a checklist [PDF] of recommended actions to take.

    "If working with Ukrainian organizations, take extra care to monitor, inspect, and isolate traffic from those organizations; closely review access controls for that traffic," added CISA, which also advised reviewing backups and disaster recovery drills.

    Continue reading

Biting the hand that feeds IT © 1998–2022