The debate over whether Chinese AI models were trained on data from US-based LLMs like Anthropic and OpenAI is secondary to the legality of how those US models were trained. If it was permissible for Anthropic and OpenAI to use the entire internet and copyrighted works for training, then similar downstream use by others should also be considered legal. This perspective suggests that content derived from such training data becomes part of the public commons, owned by everyone. AI
IMPACT Raises questions about the ethical and legal boundaries of AI training data, potentially influencing future regulations and industry practices.
RANK_REASON The item discusses the legality of AI training data usage and its implications, framed as an opinion piece.
Read on Mastodon — fosstodon.org →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →