PulseAugur
EN
LIVE 19:12:17

Reddit discussion: LLMs trained on 'stolen IP' not proprietary tech

A discussion on Reddit's r/LocalLLaMA subreddit posits that large language models (LLMs) are not built on proprietary technology but rather on vast amounts of proprietary data, which the poster characterizes as stolen intellectual property. The conversation, addressed to a 'Michael,' suggests that the underlying architecture of LLMs is open, but their training datasets are a significant point of contention regarding ownership and legality. AI

IMPACT Raises questions about the ethical and legal sourcing of data used to train large language models.

RANK_REASON The cluster consists of a single Reddit post discussing the nature of LLM training data, which falls under commentary.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Reddit discussion: LLMs trained on 'stolen IP' not proprietary tech

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/ZenaMeTepe ·

    Dear Michael, LLMs don't run on proprietary technology, but they do run on lots of proprietary data aka stolen IP

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1v3nff7/dear_michael_llms_dont_run_on_proprietary/"> <img alt="Dear Michael, LLMs don't run on proprietary technology, but they do run on lots of proprietary data aka stolen IP" src="https://preview.redd.it/6y…