I'm not sure why it would be any different from how this is treated with search engines. Both scrape massive amounts of openly available data and make it available in some form. Any training data or information that a model could potentially spit out is already available through a search engine's index.
this post was submitted on 16 Jan 2024
75 points (96.3% liked)
Privacy
31998 readers
1087 users here now
A place to discuss privacy and freedom in the digital world.
Privacy has become a very important issue in modern society, with companies and governments constantly abusing their power, more and more people are waking up to the importance of digital privacy.
In this community everyone is welcome to post links and discuss topics related to privacy.
Some Rules
- Posting a link to a website containing tracking isn't great, if contents of the website are behind a paywall maybe copy them into the post
- Don't promote proprietary software
- Try to keep things on topic
- If you have a question, please try searching for previous discussions, maybe it has already been answered
- Reposts are fine, but should have at least a couple of weeks in between so that the post can reach a new audience
- Be nice :)
Related communities
Chat rooms
-
[Matrix/Element]Dead
much thanks to @gary_host_laptop for the logo design :)
founded 5 years ago
MODERATORS
75
UK privacy watchdog to examine practice of web scraping to get training data for AI
(therecord.media)
The UK has a data protection agency? Does the UK know? Have they been asleep for the past 20 years?