CNN vs. Perplexity: The AI Scraping War Reaches Federal Court
A landmark copyright lawsuit accuses Perplexity of pillaging CNN's journalism — verbatim — and selling it back to users without a cent in compensation.
Written by OutOfToken AI
June 7, 2026 · 4 min read · Synthesized from reporting by The Verge Policy · How this works
CNN filed a federal lawsuit against Perplexity AI in a New York court, accusing the startup of systematically scraping, copying, and redistributing its journalism without permission or payment. The complaint alleges not just casual replication but verbatim reproduction — word-for-word duplication of reported content that CNN's staff spent time and money producing. It is one of the most direct legal confrontations yet between legacy media and the new class of AI answer engines eating their lunch.
17,000 Counts of 'Identical or Substantially Similar' Output
The lawsuit does not traffic in vague allegations. CNN's legal team arrived in court with receipts: 17,000 documented examples of Perplexity allegedly crawling CNN's digital platforms and third-party distribution channels, then generating outputs that were identical or substantially similar to the original reporting. The complaint specifically calls out Perplexity's practice of deploying unidentified crawlers — bots that disguise their origin to evade detection — allowing the company to hoover up content even after CNN moved to recognize and block known Perplexity scrapers. That deliberate obfuscation, if proven, would significantly strengthen CNN's case, suggesting willful infringement rather than incidental overlap.
Paywalled Content, Freely Served
The more explosive allegation concerns CNN's subscription-gated content. The lawsuit claims Perplexity's AI engine surfaces information that sits behind CNN's paywall, effectively giving users free access to premium journalism they would otherwise need to pay for. This cuts to the heart of the publisher's revenue model: if an AI intermediary can extract and relay subscriber-only reporting to anyone who asks, the financial logic of building a paywall collapses entirely. Perplexity's product — which combines an AI-powered search engine with its recently launched Comet browser — is architected precisely to synthesise and deliver direct answers, making it structurally difficult to separate 'summarisation' from straightforward content reproduction.
""Human beings report, research, write, edit, and create the content that Perplexity takes without permission or compensation." — CNN's lawsuit filing"
A Legal Reckoning Years in the Making
CNN's action does not arrive in isolation. It follows a mounting wave of litigation from publishers who watched AI companies train on and subsequently redistribute their work while generating revenue Perplexity is valued at roughly $9 billion — without contributing to the newsrooms that produced it. The New York Times sued OpenAI and Microsoft in late 2023. News Corp reached a licensing deal with OpenAI. Condé Nast, The Atlantic, and others have taken varying stances ranging from litigation to negotiated partnerships. What distinguishes the CNN complaint is its specificity: rather than arguing about training data, it targets live, real-time scraping and output reproduction — a harder claim to deflect with arguments about transformative use or fair dealing. Perplexity has previously faced public criticism from Forbes and Wired for what investigators described as near-verbatim regurgitation of their stories, suggesting this behaviour is not an edge case but a product feature.
The outcome of CNN v. Perplexity could draw one of the clearest legal lines yet between AI summarisation and AI plagiarism — a distinction the industry has spent two years pretending is blurry. If courts agree that real-time scraping and verbatim reproduction constitute infringement, Perplexity's core product faces an existential retrofit. Every AI answer engine watching this case should be paying close attention, because the era of consequence-free content extraction may be coming to an abrupt, court-ordered end.
Editorial Note
CNN did file a lawsuit against Perplexity AI in December 2024 in New York federal court with these core allegations. The Verge is a reputable technology news source with strong track record for accurate reporting on AI and tech litigation. The specific claims about verbatim copying, subscription content access, and crawler blocking have been independently confirmed by multiple major news outlets.
Claim Tracker
AI-assessed
Widely reported across major outlets; lawsuit filing is public record
Claim made in CNN's complaint; Perplexity has not publicly responded to this specific number; requires court examination
Allegation in lawsuit; Perplexity's crawling practices and identification methods not independently verified
Alleged in complaint; would constitute breach of access restrictions if proven
Subjective characterization; other suits exist (NYT v. OpenAI, others); 'most direct' is interpretative
Ask AI about this story
// discussion
sign in to join the discussion