WeSearch
Hub / Tags / Training Data
TAG · #TRAINING-DATA

Training Data coverage.

Every story in the WeSearch catalog tagged with #training-data, chronological, with view counts. Subscribe to the per-tag RSS feed to follow this topic in your reader of choice.

12 stories tagged with #training-data, in publish-time order across the WeSearch catalog. Tag pages update as new stories ingest.

⌘ RSS feed for this tag →   or   search "Training Data"

RELATED TAGS
#ai1#books1#publishing1#ethics1
TECHSPOT

AI firms are quietly buying and destroying millions of printed books to train their models

Its pitch is simple: older books offer cleaner data. "The world's best AI training data is sitting on a shelf," the company says on its website. "Books...…

3 views ·
#ai#books
ARXIV.ORG

OriginBlame: Record- and Token-Level Data Provenance for AI Training Datasets

arXiv:2607.13037v1 Announce Type: new Abstract: When a data contributor requests removal, model trainers face a practical gap: unlearning algorithms require a forget set, yet no to…

34 views ·
#originblame#record-#token-level
TECHMEME

Nvidia unveils Cosmos 3, an open physical AI foundation model, to help robots and autonomous cars better understand the real world with limited training data (Ina Fried/Axios)

Ina Fried / Axios : Nvidia unveils Cosmos 3, an open physical AI foundation model, to help robots and autonomous cars better understand the real world with limited training data — …

42 views ·
QUARTZ

AI training data is becoming a seller's market. Here's what it's worth

29 views ·
R/LOCALLLAMA

Turning every "no thats not what i meant" in chat into actual LoRA training data

32 views ·
ARXIV CS.AI

Pretraining Data Exposure in Large Language Models: A Survey of Membership Inference, Data Contamination, and Security Implications

Large Language Models (LLMs) have become the predominant paradigm in NLP, advancing both research and industry. As model sizes and pretraining data grow, concerns about Pretraining…

34 views ·
#artificial intelligence#machine learning#data privacy
ARXIV CS.AI

Brain-LLM Alignment Tracks Training Data, Not Typology

Brain-LLM alignment is well established in English, yet the brain's language network is neuroanatomically universal across languages. Does alignment also generalize cross-linguisti…

19 views ·
#language#artificial intelligence#neuroscience
R/MACHINELEARNING

LQS v3.1 — an open methodology for rating AI training data (multi-oracle consensus + signed certificates) [P]

38 views ·
ANDROID AUTHORITY

COROS thinks ChatGPT should analyze your training data

COROS’ new AI integration lets athletes digest training history, recovery, race readiness and more using ChatGPT and Claude.…

21 views ·
#fitness#wearables#ai
TECHMEME

Internal chats: in March, xAI offered employees $420 in exchange for completed tax filings as training data for Grok, but the bonuses haven't been paid out (Carmen Arroyo/Bloomberg)

33 views ·
R/MACHINELEARNING

How are you handling training data when public datasets don't match your use case? [D]

34 views ·
R/SINGULARITY

What happened to the issue of companies running out of training data for LLMs?

30 views ·