PulseAugur
EN
LIVE 15:05:48

AI company scans world's books for training data, bypassing copyright

An AI company is reportedly scanning "all the books in the world" for training data, bypassing copyright considerations. This approach is seen as an easier alternative to navigating copyright laws for data acquisition. The practice has drawn attention and criticism. AI

IMPACT Raises questions about data sourcing and copyright in AI development, potentially influencing future policy and industry practices.

RANK_REASON The cluster consists of social media posts discussing an AI company's data acquisition methods, rather than a primary announcement or research paper.

Read on Mastodon — mastodon.social →

AI-generated summary · Google Gemini · from 2 sources. How we write summaries →

AI company scans world's books for training data, bypassing copyright

COVERAGE [2]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    The AI company apparently found destructively scanning ‘all the books in the world’ easier than dealing with copyright in its quest for training data. #Books #B

    The AI company apparently found destructively scanning ‘all the books in the world’ easier than dealing with copyright in its quest for training data. #Books #BookSky #Anthropic #AI 📚📖 RE: https://bsky.app/profile/did:plc:vovinwhtulbsx4mwfw26r5ni/post/3msdjcoioue2o

  2. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    The AI company apparently found destructively scanning ‘all the books in the world’ easier than dealing with copyright in its quest for training data. #Books #B

    The AI company apparently found destructively scanning ‘all the books in the world’ easier than dealing with copyright in its quest for training data. #Books #BookSky #Anthropic #AI 📚📖 RE: https://bsky.app/profile/did:plc:vovinwhtulbsx4mwfw26r5ni/post/3msdjcoioue2o