PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 19, 20260 citationsOpen Access

Data Integrity as the Terminal Constraint in AI-Driven Advertising: An Information-Theoretic Analysis of Conversion Fraud and Agentic Threat Evolution

View Full Paper
IIIgor IvitskiyDSDmytro SavchenkoDSDana Sydorenko

Key Points

  • The analysis aims to investigate how adversarial contamination of conversion data affects AI advertising performance.
  • Developed an information-theoretic framework to analyze conversion fraud and data integrity.
  • Created a four-generation taxonomy of advertising fraud evolution.
  • Applied PAC learning theory to assess algorithmic performance under fraud.
  • Conducted an economic analysis of the impacts of conversion data corruption.
  • Demonstrated that data poisoning severely limits algorithmic advertising performance.
  • Proved that increasing fraud rates inflate sample complexity by 4x to over 100x, hindering learnable signals.
  • Estimated total losses from conversion fraud extend beyond $84 billion annually, projected to reach $172 billion by 2028.
  • Identified a feedback loop where increasing budget allocations to fraudulent sources result in distorted return reports.

Abstract

Modern AI-driven advertising systems, including Google Performance Max and Meta Advantage+, rely on conversion event data as their primary training signal. This paper demonstrates that adversarial contamination of these signals constitutes the binding constraint on algorithmic advertising performance, a constraint that no amount of computational sophistication can overcome. We develop an information-theoretic framework showing that the data processing inequality imposes a hard ceiling on optimization quality: when conversion data is poisoned, algorithmic decisions inherit and amplify the corruption. We formalize a four-generation taxonomy of advertising fraud evolution, from volumetric bot traffic (Generation 1) through sophisticated invalid traffic (Generation 2), conversion-level data poisoning (Generation 3), to an emerging class of autonomous agent collusion (Generation 4). Applying results from PAC learning theory, we prove that at empirically observed fraud rates of 25% in lead generation (rising to 67-85% in audience networks), sample complexity inflates by 4x to over 100x, pushing many campaigns past the threshold of learnable signal. Economic analysis reveals that the total cost extends well beyond the estimated 84 billion in direct annual losses (projected to 172 billion by 2028): corrupted optimization data creates a compounding feedback loop in which platforms allocate increasing budget to fraudulent sources, producing a measurable divergence between reported and verified returns. We propose signal hygiene, a verification paradigm that validates the event rather than surveilling the user, as the structurally privacy-compatible alternative to surveillance-based fraud detection.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Ivitskiy et al. (2026) studied this question.

synapsesocial.com/papers/6996a8b5ecb39a600b3efbc1https://doi.org/10.5281/zenodo.18675361
Ask AI
Helpful
Bookmark
Share
View Full Paper