PulseAugur
EN
LIVE 18:40:46

US LLM training data legality questioned amid China AI model debate

The debate over whether Chinese AI models were trained on data from US-based LLMs like Anthropic and OpenAI is secondary to the legality of how those US models were trained. If it was permissible for Anthropic and OpenAI to use the entire internet and copyrighted works for training, then similar downstream use by others should also be considered legal. This perspective suggests that content derived from such training data becomes part of the public commons, owned by everyone. AI

IMPACT Raises questions about the ethical and legal boundaries of AI training data, potentially influencing future regulations and industry practices.

RANK_REASON The item discusses the legality of AI training data usage and its implications, framed as an opinion piece.

Read on Mastodon — fosstodon.org →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

US LLM training data legality questioned amid China AI model debate

COVERAGE [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    IMHO? It doesn't matter if the Chinese used USA LLMs to train their own or not. If it was legal for Anthropic and OpenAI to download the entire Internet and sca

    IMHO? It doesn't matter if the Chinese used USA LLMs to train their own or not. If it was legal for Anthropic and OpenAI to download the entire Internet and scan copyrighted works to train their models? Then it's legal to do a downstream version. If something is based on the # co…