Skip to main content

Ethical and Social Challenges with developing Automated Methods to Detect and Warn potential victims of Mass-marketing Fraud (MMF)

Monica Whitty ; Matthew Edwards ; Michael Levi ; Claudia Peersman ; Awais Rashid ; Angela Sasse ; Tom Sorell ; Gianluca Stringhini (2017) — Proceedings of the 26th International Conference on World Wide Web Companion - WWW '17 Companion

Citation activity

Total citations · Google Scholar
Unavailable
Average per calendar year
Citation count unavailable

Counts reflect Google Scholar’s coverage and do not measure research quality. Source, calculation and limitations

What do these research terms mean?
Preprint
A manuscript shared before formal peer review and publication. Check whether a later published version is available.
Dataset
A collection of data or examples for others to inspect or reuse. It can appear in Library search and topic mapping, but RSRC does not use dataset records as evidence in Research Insights.
Dissertation or thesis
Research submitted for an academic degree. This describes its format, not its reliability.
Journal article
An article published in a journal. This label alone does not establish peer review, study quality, or how well the findings apply elsewhere.
Qualitative research
Examines experiences, meanings, or processes, often through interviews or observations. It can explain how something happens without estimating how common it is.
Quantitative research
Uses numerical measurements to describe patterns or test relationships. A relationship between two measurements does not by itself show that one causes the other.
Systematic review
Uses a planned, documented method to find and assess research addressing a question. Its conclusions still depend on the included studies and what the search covered.
Meta-analysis
Statistically combines results from multiple studies. Combining studies does not remove weaknesses in their design or make unlike populations interchangeable.
Not classified
This record has no recognized label in this filter. It does not mean the publication used no method, or that no research exists.

Definitions draw on DataCite resource types; Cochrane review methods; NLM: association and causation. RSRC’s dataset and classification rules are explained in our methodology.

Citation tools


              
              

Transparency

Evidence and review status

This page contains AI-generated content. No human content review or subject-matter-expert review is recorded.

Source basis
Downloaded PDF
Source updates
No notice found at last check
AI-generated page content
Yes
Automated checks
Passed
Administrative approval
Yes
Human content review
Not recorded
Subject-matter-expert review
Not recorded
How this was prepared
Source basis

RSRC downloaded and privately stored a copy of the paper for internal analysis. The PDF is not offered to viewers from this page.

  • PDF added to RSRC:
Source updates

No incoming update notice was found in the dated Crossref response. Coverage is incomplete, particularly for corrections and expressions of concern; this is not a guarantee that the source is valid or unchanged.

  • Last source-status attempt:
AI-generated page content

AI-generated research notes displayed on this page: Synopsis, Identified gaps, Methods, Limitations, Future work. The paper itself is not described as AI-generated.

  • Document analysis recorded:
  • Synopsis generation recorded:
  • Page record updated:
Automated checks

The current, source-bound synopsis passed the recorded versioned publication checks.

  • Checks completed:
View passed checks (3)
  • Length, completeness, repetition, refusal, boilerplate, and active-markup screening
  • Numerical claims checked against the available source text
  • English-source lexical grounding check
Administrative approval

An authenticated administrator approved the bibliographic record for public Library display. This is not a review of every research claim.

  • Approved for public display:
Human content review

No human review is recorded for the AI-generated content displayed on this page.

Subject-matter-expert review

RSRC has not recorded review of this content by a subject-matter or methods expert.

Review-state definitions
Found a possible error? Request a correction.

Synopsis

The publication presents an interdisciplinary effort to develop automated methods for detecting mass-marketing fraud (MMF) and warning potential victims, with a focus on mass-market scams including romance scams. Its central aim is to prevent victimization by enabling early identification of deceptive communication and grooming strategies across multiple online channels. The authors describe an ongoing project that blends psychology, media studies, criminology, linguistics, and human-computer interaction to identify signals of deception, grooming, and persuasive requests within end-user communications and profiles. The paper outlines a research program that would use personal data from dating sites, employment pages, emails, and other sources to train and test machine-learning and pattern-recognition approaches. Proposed methods include supervised learning (e.g., random forests) and clustering to distinguish scam-related behavior and to uncover linguistic and stylistic indicators typical of victims and scammers. They also consider analyzing sociotechnical features such as response patterns, profile descriptions, and multi-channel interactions. Ethical and social challenges are given substantial attention: data anonymization, informed consent, potential false positives, and the need to balance protection with user autonomy. The authors acknowledge the difficulty of making automated judgments about authenticity and the risk of harming the very individuals they aim to help, suggesting human-in-the-loop or warning-message strategies as possible mitigations. The discussed implications emphasize that while automated MMF detection could reduce financial and psychological harm, it raises privacy concerns and governance questions about how such systems are deployed, how risks of misclassification are managed, and how user trust is maintained. The publication stops short of presenting empirical results, instead outlining the conceptual framework, methodological directions, and ethical constraints guiding future work.

Identified Gaps

Existing scam education focused on idealized individual behavior and knowledge may not prevent victimization; awareness can coexist with vulnerability. MMF detection is difficult because offenders use personalized, adaptive, long-term communication across multiple channels. The paper also identifies an unresolved need to design detection and warning systems that balance prevention with privacy, autonomy, false-positive harms, and trust.

Methods

This extended abstract describes early planning for an interdisciplinary MMF detection project. The proposed research analyzes dating and employment profiles and communications from victims, scammers, and noncriminal contacts. It plans to examine psychological, linguistic, behavioral, and socio-technical indicators, using anonymization, supervised machine learning (including random forests), and profile clustering. If detection proves effective, the team plans a browser-extension proof of concept and HCI-informed tests of warning messages.

Limitations

The paper reports planned work rather than completed empirical findings, validation results, or performance measures. Detection may produce false positives that harm genuine users and relationships. The proposed system would require analysis of highly personal data and may make authenticity judgments without the other person's knowledge. Victims may also distrust or disregard warnings, limiting intervention effectiveness.

Future Work

Develop and evaluate a proof-of-concept browser extension that detects suspected scam accounts and warns users. Test the effectiveness of alternative warning-message designs. Explore warning approaches that retain human judgment, such as guided authenticity checks, and consider whether tools should be situated on dating sites or support networks.

See how this publication connects to RSRC's living evidence syntheses through current citations and research-topic mapping.