Vivold Consulting

NYT escalates legal battle over AI training data by targeting Perplexity's use of copyrighted material

Key Insights

The New York Times filed a lawsuit against Perplexity alleging unauthorized use of copyrighted reporting in training and generation. The case intensifies the legal fight over publisher rights, AI model training, and acceptable use boundaries.

Stay Updated

Get the latest insights delivered to your inbox

Copyright law hits its AI inflection point


The New York Times' lawsuit against Perplexity raises key questions about how AI companies source, store, and transform content. While numerous publishers have negotiated licensing agreements with AI labs, others are choosing to litigatein hopes of shaping the rules of the ecosystem.

What the lawsuit claims


Though details will evolve in court, the allegations reflect broader industry tensions:
- Perplexity allegedly used Times content without a license for model training and outputs.
- The Times argues that AI-generated paraphrasing still represents derivative use.
- The suit seeks to clarify whether AI companies must pay for both training and downstream generation rights.

Why this matters for model builders


Legal exposure is rising:
- Companies will need strong documentation of training data sources.
- Licenses may expand to include synthetic variants of copyrighted text.
- Publishers may push for per-query monetization in addition to flat training fees.

A market moving toward frameworks


This caseand others like itmay catalyze industry-wide standards for ethically sourced and legally defensible datasets. It could also accelerate a shift toward publisher consortiums selling structured, AI-ready archives.

Related Articles

Discovery Loop aims to automate science itself - and Google is funding the startup draining its own bench, as Hassabis exits the DeepMind CEO role

Jeff Dean, Google's chief scientist and 30th employee, is leaving after 27 years to found Discovery Loop, a public benefit corporation using AI to automate scientific research - taking co-founders Sanjay Ghemawat, Quoc Le (Google Brain), and Oriol Vinyals (DeepMind) with him. Google is a founding investor and cloud partner, supplying compute for at least the first year, with Radical Ventures and Khosla Ventures co-leading the seed. In the same announcement, Demis Hassabis steps down as DeepMind CEO to become chairman and Alphabet chief scientist, with Koray Kavukcuoglu taking over Gemini model development. Alphabet stock fell about 4%.

Abbott orders audits of every new project as ERCOT's queue hits 474GW, roughly 90% of it data centres

Governor Greg Abbott announced that all new Texas data-centre projects must be audited by the Public Utility Commission and grid operator ERCOT - a sharp turn for a state whose loose regulation and cheap power made it second only to Virginia for data centres. The trigger is a staggering queue: ERCOT's interconnection requests doubled from 233GW in January to 474GW, about 90% data centres, more than five times the grid's all-time peak demand. Audits will demand power and water use, noise mitigation, light controls, tax-incentive use, and ownership details - after a voluntary survey that most operators simply ignored.

Volta and Bitdeer will build a 133MW Nvidia Vera Rubin data centre in Norway - Anthropic's latest move in a compute land grab

Anthropic has reportedly signed a $10 billion, six-year compute deal with Volta, an AI cloud startup founded only earlier this year, per Bloomberg. Volta is partnering with crypto-mining firm Bitdeer to develop the data centre - located in Norway, delivering 133 megawatts, and running Nvidia's Vera Rubin architecture - and is a member of Nvidia's Cloud Partner programme. It caps an aggressive capacity spree that also includes recent compute deals with SpaceX and Amazon, as Anthropic races rivals for the scarcest input in the industry.