• danc4498@lemmy.world
    link
    fedilink
    English
    arrow-up
    41
    ·
    4 hours ago

    Why would AI want a continuous deal with Reddit? Don’t they get all the data they need the first time? I doubt the new content is worth as much as the previous deal… maybe I don’t understand what these deals are for.

    • UnderpantsWeevil@lemmy.world
      link
      fedilink
      English
      arrow-up
      26
      arrow-down
      1
      ·
      edit-2
      4 hours ago

      Half the joke is that Reddit was ground zero for AI slop even before AI had gone mainstream.

      The company got harvested back before the AI firms were overly worried with cross-contamination.

    • XLE@piefed.social
      link
      fedilink
      English
      arrow-up
      5
      ·
      3 hours ago

      Presuming they took all the data, a one-time deal would only be good if knowledge gathering actually stopped after the cutoff year- but for recent things like tech and news, the models have to keep learning and adding to their repositories.

      The returns on that value sharply diminish, of course, but I think they’re still necessary. Which will leave everybody in a bind that is very funny.

        • ryper@lemmy.ca
          link
          fedilink
          English
          arrow-up
          4
          ·
          2 hours ago

          The API restrictions and login requirements are meant to make scraping hard enough to make a deal worthwhile.

        • XLE@piefed.social
          link
          fedilink
          English
          arrow-up
          1
          ·
          2 hours ago

          Probably because Reddit has lawyers, and money, and a little willingness to lock down their content. Unlike individual creators, they can actually file a lawsuit