PulseAugur
EN
LIVE 14:41:33

AI data extraction needs collective redistribution, not individual pay; open-weight models urged

The current system for distributing remuneration to rights holders is insufficient for the vast amount of data used to train AI models, which includes not only professional content but also public domain works and user-generated content. A more appropriate response to AI data extraction is collective redistribution to support public information infrastructure, rather than individual compensation, due to the difficulty in tracing specific creators. Additionally, there is a call for American AI labs to release more frontier-grade open-weight models under permissive licenses to foster startup innovation, with examples like NVIDIA's Nemotron and Thinking Machines' Inkling being cited as positive steps, though many leading models remain proprietary. AI

IMPACT Calls for collective redistribution of AI training data revenue and the release of more open-weight models could shape future AI development and accessibility.

RANK_REASON The cluster discusses policy implications and calls for action regarding AI data and model releases, rather than reporting on a specific event.

Read on Mastodon — sigmoid.social →

AI-generated summary · Google Gemini · from 2 sources. How we write summaries →

AI data extraction needs collective redistribution, not individual pay; open-weight models urged

COVERAGE [2]

  1. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    "Existing collective management infrastructure is designed to distribute remuneration to professional rights-holders. That is a legitimate function, but an insu

    "Existing collective management infrastructure is designed to distribute remuneration to professional rights-holders. That is a legitimate function, but an insufficient one. The training of foundation models has not drawn exclusively on the professional content economy. It has dr…

  2. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    "American labs need to release frontier-grade open-weight models under licenses that startups can actually build on. There has been progress. NVIDIA’s Nemotron

    "American labs need to release frontier-grade open-weight models under licenses that startups can actually build on. There has been progress. NVIDIA’s Nemotron models are commercially usable under NVIDIA’s own permissive license. Thinking Machines released Inkling under Apache 2.…