Tag

Fair Use

All articles tagged with #fair use

Third Circuit Rules AI Training on Competitor's Legal Data Violates Copyright
technology8 days ago

Third Circuit Rules AI Training on Competitor's Legal Data Violates Copyright

The Third Circuit Court of Appeals ruled that training an AI system on a competitor's copyrighted editorial summaries constitutes infringement, not fair use. The court affirmed that Westlaw headnotes are protected works and that ROSS Intelligence’s use of them to build a rival legal research tool was minimally transformative. This marks the first federal appellate decision addressing fair use in AI training, though the full opinion remains under seal.

Third Circuit rules AI training on copyrighted legal data is not fair use
technology9 days ago

Third Circuit rules AI training on copyrighted legal data is not fair use

The Third Circuit Court of Appeals has affirmed a ruling against ROSS Intelligence, determining that using copyrighted legal summaries to train a competing AI product constitutes copyright infringement rather than fair use. The court found that Thomson Reuters’ Westlaw headnotes possess sufficient 'creative spark' to be protected, and that ROSS’s use was not transformative enough to justify the copying. This decision marks the first federal appellate ruling on fair use in AI training, though the full opinion remains under seal pending redaction reviews.

Internal Memos Warn AI Training on News Could Be the Largest Theft of Labor
technology22 days ago

Internal Memos Warn AI Training on News Could Be the Largest Theft of Labor

Newly unsealed court documents reveal Microsoft and OpenAI expressed serious concerns that training AI models on millions of news articles could amount to the largest theft of labor in history, potentially undermining paywalled publishing and the quality of their language models; the companies argue the training falls under fair use and is transformative, while publishers seek licensing and stronger protections as the case proceeds.

technology1 month ago

Authors accuse OpenAI of training GPT on pirated books, seek summary judgment

A group of authors filed a motion for summary judgment in New York accusing OpenAI of training its GPT models on pirated LibGen copies, concealing the piracy and building the business on mass piracy, and asking the court to find liability and that such copying isn’t fair use; the filing covers 194 titles and also argues Microsoft is vicariously liable, while OpenAI counters with a cross-motion asserting fair use as part of broader lawsuits against OpenAI and Microsoft.

US DOJ Sides with OpenAI in High-Stakes Copyright Fight Over AI Training Data
technology1 month ago

US DOJ Sides with OpenAI in High-Stakes Copyright Fight Over AI Training Data

The U.S. Department of Justice urged a New York court to let OpenAI broadly access copyrighted material for training its AI, arguing that such use is fair, transformative, and essential to maintaining U.S. leadership in AI. The Intercept and other media plaintiffs contend this would amount to an uncompensated transfer of IP rights to tech companies. The case, which involves open DMCA claims against OpenAI and has been consolidated with other media lawsuits (including The New York Times, Tribune, and Ziff Davis), has seen mixed rulings and continues as arguments over fair use and the impact on press freedom unfold.

US Government Backs OpenAI in NYT AI Copyright Case
technology1 month ago

US Government Backs OpenAI in NYT AI Copyright Case

The Trump Administration filed a DOJ brief backing OpenAI in The New York Times’ copyright lawsuit, arguing that training AI on copyrighted articles is extraordinarily transformative and likely fair use, a stance the government says could safeguard AI leadership and innovation in the U.S. while not binding on the judge; the filing comes amid broader AI copyright battles and drew mixed reactions from The Times, the Authors Guild, and others.

US backs OpenAI in NYT copyright case, arguing fair use covers AI training
ai1 month ago

US backs OpenAI in NYT copyright case, arguing fair use covers AI training

The Trump administration filed a statement of interest in The New York Times’ copyright lawsuit against OpenAI, arguing that training AI models on copyrighted text qualifies as fair use and should not require licensing, a position the DOJ says would support scientific progress and American prosperity and could set a precedent for future AI licensing disputes.

Artists push back on AI training, scoring landmark wins and guardrails
technology2 months ago

Artists push back on AI training, scoring landmark wins and guardrails

A wave of lawsuits by authors, musicians and visual artists challenges AI systems trained on copyrighted works, arguing copyright infringement or terms-of-service violations. Notable developments include Andrea Bartz’s suit against Anthropic resulting in a $1.5 billion settlement and a ruling that using pirated ebooks for training can be fair use in certain contexts, ongoing cases against Meta, Google, Suno and others, and class actions from illustrators like Sarah Andersen. While plaintiffs see these cases as paving guardrails, many creators warn that the livelihoods of independent artists remain at risk as tech companies push for broad data-use rights through their platforms and terms of service.

Anthropic's $1.5B Settlement Marks Major Win in Copyright Case Over AI Training Data
technology2 months ago

Anthropic's $1.5B Settlement Marks Major Win in Copyright Case Over AI Training Data

A San Francisco federal judge approved a $1.5 billion settlement between Anthropic and authors who accused the company of using pirated books to train its AI. Described as the largest copyright class-action settlement, the deal pays roughly $3,000 per book for claims covering around 440,000 titles; Anthropic must delete the pirated files. The accord resolves liability for past data acquisition, not for future AI outputs or new claims.

Suno’s training data exposed: millions of songs scraped for AI music
ai2 months ago

Suno’s training data exposed: millions of songs scraped for AI music

Hacked files obtained by 404 Media indicate Suno trained its AI music generator by scraping millions of songs and metadata from YouTube Music, Deezer, and Genius, with leaked code showing scraping tools and a cappella searches; the disclosures come amid lawsuits over Suno’s training data and fair-use claims, and the breach also exposed some customer info, though Suno says the incident involved outdated code and did not expose full credit card data.

AI Giants Confront the Internet’s Distillation Dilemma
technology3 months ago

AI Giants Confront the Internet’s Distillation Dilemma

Anthropic, OpenAI, and Google warn that distillation—using outputs from one AI model to improve another—could let rivals replicate top-tier AI at a fraction of the cost, mirroring how the broader internet treats data: scrape first, justify later. While framed as a cybersecurity issue by the companies, critics argue the practice blurs legal lines and raises site costs, and the industry’s cat‑and‑mouse dynamic suggests distillation is becoming a new normal on the web.

Meta triumphs in AI copyright lawsuit, marking a significant legal victory
law1 year ago

Meta triumphs in AI copyright lawsuit, marking a significant legal victory

Meta won a legal case against authors who claimed the company used their copyrighted works without permission to train AI, with the judge ruling that the authors failed to prove their market was diluted, though the decision leaves open questions about the legality of AI training practices under fair use. The case highlights ongoing tensions between AI development and copyright law.