this post was submitted on 08 May 2024
1700 points (99.2% liked)

Technology

59438 readers
3444 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 1 year ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] Hypx@fedia.io 66 points 6 months ago (29 children)

Eventually, we will need a fediverse version of StackOverflow, Quora, etc.

[–] thfi@discuss.tchncs.de 76 points 6 months ago (12 children)

Those would be harvested to train LLMs even without asking first. 😐

[–] chameleon@kbin.social 7 points 6 months ago

SO already was. Not even harvested as much as handed to them. Periodic data dumps and a general forced commitment to open information were a big part of the reason they won out over other sites that used to compete with them. SO most likely wouldn't have existed if Experts Exchange didn't paywall their entire site.

As with everything else, AI companies believe their training data operates under fair use, so they will discard the CC-SA-4.0 license requirements regardless of whether this deal exists. (And if a court ever finds it's not fair use, they are so many layers of fucked that this situation won't even register.)

load more comments (11 replies)
load more comments (27 replies)