PulseAugur
EN
LIVE 08:33:32

AI training data could use 'trust factor' metadata

A discussion on Reddit explores the concept of incorporating a "trust factor" into AI training data. The proposed metadata would assign a score from 0.0 to 1.0, potentially decreasing for older data or data generated after 2022 due to AI's influence. This system could allow for the use of broader internet data by automatically filtering based on calculated trustworthiness, considering factors like data source, age, and reference count. AI

IMPACT This concept could enable more efficient and broader use of internet data for AI training by automating trustworthiness assessments.

RANK_REASON The item is a user discussion on a subreddit about a potential technical approach to AI training data.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

AI training data could use 'trust factor' metadata

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/freehuntx ·

    Trustfactor in training data?

    <!-- SC_OFF --><div class="md"><p>Would it be possible and make sense to add metadata to training data e.g. a trustfactor (0.0 - 1.0)? </p> <p>For example: the older data is the less trustworthy it is. And data after 2022 gets less trustworthy over time (because of ai). </p> <p>W…